Video editing method, apparatus, device, and medium

The video editing method simplifies online video collaboration by allowing users to record and edit videos using their devices, synchronizing playback, and utilizing editing templates for efficient video integration and editing.

JP7771377B2Active Publication Date: 2025-11-17BEIJING ZITIAO NETWORK TECH CO LTD
View PDF 2 Cites 0 Cited by

Patent Information

Application Number
JP2024519950
Authority / Receiving Office
JP · JP
Patent Type
Patents
Current Assignee / Owner
Priority Date
2023-03-30
Filing Date
2023-12-06
Publication Date
2025-11-17
Estimated Expiration
2043-12-06

AI Technical Summary

Technical Problem

Existing online video collaboration methods require high production costs and are inconvenient due to the need for multiple users to use multiple devices and software tools simultaneously, making video editing difficult and time-consuming.

Method used

A video editing method that allows multiple users to capture and record videos using their own devices, synchronizes playback of a shared video, and uses editing templates to easily combine and edit the recorded videos into a collaborative video.

Benefits of technology

Reduces the cost and complexity of online video collaboration by enabling easy and fast editing of videos recorded by multiple users, allowing seamless integration and synchronization of video tracks.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure 0007771377000001
    Figure 0007771377000001
  • Figure 0007771377000002
    Figure 0007771377000002
  • Figure 0007771377000003
    Figure 0007771377000003
Patent Text Reader

Abstract

The present disclosure relates to a video editing method, an apparatus, a device and a medium, the method including: obtaining a plurality of target videos, the plurality of target videos including first videos respectively shot and recorded by different user devices for the same recording task; obtaining an editing template corresponding to the plurality of target videos; the editing template including display position information of the plurality of target videos and track information of the plurality of target videos; displaying a video editing interface according to the plurality of target videos and the editing template, the video editing interface including a preview playback area and an editing track area, the editing track area including a plurality of video editing tracks, the images of the plurality of target videos are respectively displayed in the display area indicated by the respective display position information of the plurality of target videos in the preview playback area, and each target video in the plurality of target videos forms a video track clip on the video editing track in the plurality of video editing tracks respectively according to the track information. The present disclosure can effectively reduce the cost of online video collaboration.
Need to check novelty before this filing date? Find Prior Art

Description

[Technical Field]

[0001] The present disclosure relates to the technical field of video processing, and more particularly to video editing methods, apparatus, devices, and media.

[0002] Related Applications This application claims priority to Chinese invention patent application No. 202310332310.2, filed on March 30, 2023, for the invention "Video editing method, apparatus, device and medium," the entire contents of which are incorporated herein by reference. [Background technology]

[0003] Currently, as more and more users are dissatisfied with traditional video production methods, online video collaboration methods are gradually emerging. For example, same-screen game recording, online chorus, and anchor pairing are examples of online video collaboration. Online video collaboration does not require multiple users to gather in the same place to film, but it has high production costs and is less convenient. Summary of the Invention

[0004] To solve at least some or all of the above technical problems, the present disclosure provides a video editing method, apparatus, device, and medium.

[0005] An embodiment of the present disclosure provides a video editing method, the method including: acquiring a plurality of target videos, the plurality of target videos including first videos respectively captured and recorded by different user devices for the same recording task; acquiring an editing template corresponding to the plurality of target videos, the editing template including display position information of the plurality of target videos and track information of the plurality of target videos; and displaying a video editing interface based on the plurality of target videos and the editing template, wherein the video editing interface includes a preview playback area and an editing track area, the editing track area including a plurality of video editing tracks, images of the plurality of target videos are respectively displayed in display areas indicated by the respective display position information of the plurality of target videos in the preview playback area, and each target video among the plurality of target videos forms a video track clip on a video editing track among the plurality of video editing tracks based on the track information, and the timeline positions of the video track clips of the plurality of target videos partially or fully overlap.

[0006] Optionally, the target video further includes a third video obtained based on a second video, the second video being a video played back during the process of being filmed and recorded by the different user device for the recording task, and the third video being used to represent the playback state of the second video during the filming and recording process.

[0007] Optionally, the different user devices include a target user device, and a video operation control in a triggerable state is displayed on a user interface provided by the target user device for the recording task, and the video operation control in a triggerable state is not displayed on a user interface provided by another user device other than the target user device among the different user devices for the recording task, and when the video operation control is in a triggerable state, in response to detecting a user trigger operation on the video operation control on the target user device, the playback state of the second video is adjusted based on the user trigger operation.

[0008] Optionally, the different user devices include a first user device and a second user device, the first user device being a device that introduces the second video into the recording task, and when the recording task is not closed and the first user device has not finished the recording task, the target user device is the first user device, and when the recording task is not closed and the first user device has finished the recording task, the target user device is the second user device.

[0009] Optionally, the video editing track corresponding to the tertiary video is a main track, and the video editing track corresponding to the primary video is a picture-in-picture track.

[0010] Optionally, when the plurality of target videos includes only a first video, the video editing track corresponding to the first video shot by a third user device among the different user devices is a main track, and the video editing track corresponding to the first video shot by a user device other than the third user device in the plurality of target videos is a picture-in-picture track, wherein the third user device is a device that initiates the recording task.

[0011] Selectably, in response to receiving an editing request from a fourth user device among the different user devices, displaying a template selection page, wherein the template selection page includes a plurality of video layout effect diagrams, the video layout effect diagrams including a plurality of regions, each region among the plurality of regions corresponding to a video display position and a video editing track; in response to detecting a selection operation on a target effect diagram in the plurality of video layout effect diagrams, determining an editing template corresponding to the target effect diagram as an editing template corresponding to the plurality of target videos.

[0012] Selectively, determining editing templates corresponding to the plurality of target videos based on editing templates corresponding to the target effect diagrams includes displaying layout preview images of the plurality of target videos based on a target effect diagram selected by the target user from the plurality of video layout effect diagrams and the plurality of target videos, and determining editing templates corresponding to the plurality of target videos based on an adjustment operation by the target user on a display position of the target video in the layout preview image.

[0013] Optionally, before acquiring the multiple recorded videos, the method further includes the steps of: creating a recording task in response to receiving a collaborative shooting request from the target user device, and displaying a user interface corresponding to the recording task on the target user device, wherein the user interface displays an add video control and a user invitation control; acquiring video information uploaded by the target user device in response to the add video control being triggered, and obtaining a second video based on the video information, wherein the video information includes a local video file and / or a network video link; and generating invitation information for the recording task in response to the user invitation control being triggered, and the target user device sending the invitation information to a designated user device, wherein the invitation information is used to invite the designated user device to participate in the recording task.

[0014] Selectably, a first area and a second area are displayed on the user interface of both the target user device and the other user device performing the recording task, wherein the first area is used to display an image screen of a first video captured and recorded by each of the user devices, and the second area is used to display an image screen of the second video.

[0015] Optionally, a start recording control is further displayed on the user interface of the target user device, and the target user device triggers the add video control, and the start recording control is in a non-triggerable state when it is not detected that the target user device triggers the user invitation control, wherein a start recording control in a non-triggerable state is not triggered to record video, and when it is detected that the target user device triggers the add video control and / or the target user device triggers the user invitation control, the start recording control is in a triggerable state, wherein a start recording control in a triggerable state is triggered to start video recording.

[0016] Optionally, the step of acquiring the multiple recorded videos includes a step of recording and obtaining a first video based on frame images collected by a front camera of the user device executing the recording task in response to a trigger of a recording start control in a triggerable state, and when a second video is acquired by the target user device, the user device executing the recording task synchronously plays back the second video during the video recording process, and records the playback process of the second video to obtain a third video.

[0017] An embodiment of the present disclosure further provides a video editing device, including: a video acquisition module for acquiring a plurality of target videos, the plurality of target videos including first videos each captured and recorded by a different user device for a same recording task; a template acquisition module for acquiring editing templates corresponding to the plurality of target videos, the editing templates including display position information of the plurality of target videos and track information of the plurality of target videos; and a video editing module for displaying a video editing interface based on the plurality of target videos and the editing template, wherein the video editing interface includes a preview playback area and an editing track area, the editing track area including a plurality of video editing tracks, images of the plurality of target videos are respectively displayed in display areas indicated by the display position information of the plurality of target videos in the preview playback area, and each target video in the plurality of target videos forms a video track clip on a video editing track in the plurality of video editing tracks based on the track information, and the timeline positions of the video track clips of the plurality of target videos partially or fully overlap.

[0018] An embodiment of the present disclosure further provides an electronic device, the electronic device including a processor and a memory for storing executable instructions by the processor, the processor being used to read the executable instructions from the memory and execute the instructions to implement the video editing method provided by the embodiment of the present disclosure.

[0019] An embodiment of the present disclosure further provides a computer-readable storage medium having a computer program stored therein, the computer program being used to implement the video editing method provided by the embodiment of the present disclosure.

[0020] According to the above technical solution provided by the embodiments of the present disclosure, multiple target videos (including first videos shot and recorded by different user devices for the same recording task) are obtained, and editing templates corresponding to the multiple target videos (including display position information of the multiple target videos and track information of the multiple target videos) are obtained. Then, a video editing interface can be directly displayed based on the multiple target videos and the editing templates. The video editing interface includes a preview playback area and an editing track area, where the editing track area includes multiple video editing tracks. Images of the multiple target videos are displayed in display areas indicated by the display position information of each of the multiple target videos in the preview playback area. Each target video among the multiple target videos forms a video track clip on one of the multiple video editing tracks based on the track information, and the timeline positions of the video track clips of the multiple target videos partially or fully overlap. This aspect allows videos shot and recorded by multiple users for the same recording task to be easily and quickly edited, effectively reducing the cost of online video collaboration.

[0021] It should be understood that the contents of this section are not intended to identify key or important features of the embodiments of the present disclosure, nor are they intended to limit the scope of the present disclosure. Other features of the present disclosure will be readily apparent from the following description. [Brief explanation of the drawings]

[0022] The accompanying drawings herein, which are incorporated herein as part of this specification, illustrate suitable embodiments of the present disclosure and, together with the description, are used to explain the principles of the present disclosure.

[0023] In order to more clearly describe the technical solutions in the embodiments of the present disclosure or the prior art, the following briefly describes the drawings that need to be used in the description of the embodiments or the prior art, and obviously, those skilled in the art can obtain other drawings based on these drawings without any creative work. [Figure 1] 1 is a schematic flowchart of a video editing method provided by an embodiment of the present disclosure. [Figure 2] FIG. 2 is a schematic diagram of a user interface provided by an embodiment of the present disclosure. [Figure 3] FIG. 2 is a schematic diagram of a user interface provided by an embodiment of the present disclosure. [Figure 4] 1 is a schematic diagram of several layout effects provided by embodiments of the present disclosure. [Figure 5] 1 is a flowchart of a video editing method provided by an embodiment of the present disclosure. [Figure 6] 1 is a structural schematic diagram of a video editing device provided by an embodiment of the present disclosure; [Figure 7] 1 is a structural schematic diagram of an electronic device provided by an embodiment of the present disclosure; DETAILED DESCRIPTION OF THE INVENTION

[0024] In order to more clearly describe the above objectives, features, and advantages of the present disclosure, the solutions of the present disclosure are further described below. It should be noted that, unless mutually contradictory, the embodiments and features in the embodiments of the present disclosure can be combined with each other.

[0025] In the following description, numerous specific details are set forth in order to provide a thorough understanding of the present disclosure; however, the present disclosure may be embodied in other forms different from the embodiments herein, and it is apparent that the examples in this specification are merely some examples of the present disclosure, but not all examples.

[0026] Through research, the inventors have found that existing approaches to online video collaboration require high production costs. For example, when multiple users need to record each other's online communication performances or when different users need to record the same video viewing performances, each user needs to use at least two devices (e.g., a mobile phone and a computer turned on at the same time) to record, screencast, and perform multiple online interactions, and multiple software tools need to be turned on at the same time to achieve related functions. This not only requires time and effort, but also ultimately requires the final output to be the video output from the recording tool, making further editing difficult. Furthermore, when each user directly records their video and then combines multiple videos, each user needs to send the video to a designated user after recording, who then combines and edits the videos of multiple users, which also requires time, effort, and high costs. To solve at least one of the above problems, embodiments of the present disclosure provide a video editing method, apparatus, device, and medium, which will be described in detail below.

[0027] 1 is a schematic flowchart of a video editing method provided by an embodiment of the present disclosure, which can be performed by a video editing device. The device can be realized by software and / or hardware and generally integrated into electronic equipment. As shown in FIG. 1, the method mainly includes the following steps S102 to S106.

[0028] Step S102: Obtain a plurality of target videos, where the plurality of target videos include first videos respectively captured and recorded by different user devices for the same recording task.

[0029] The user device may be a mobile phone, a computer, a wearable device, etc., and is not limited thereto. In practice, the number of user devices may be multiple, and each user may use only one user device. When multiple users perform the same recording task, they may each use their own device to capture and record a first video corresponding to each device. In the embodiment of the present disclosure, the recording task is not limited. For example, the recording task may be expressed by capturing a user using the front camera of the user device. For example, the recording task may be expressed during the user's online interaction. For example, the recording task may be capturing a scene where the user is located using the user device. Specifically, the recording task can be flexibly set as needed, and is not particularly limited in the embodiment of the present disclosure.

[0030] In practice, the target video may further include, in addition to the first video, for example, a third video obtained based on the second video, where the second video may be a video played by a different user device during the shooting and recording process for a recording task, and the third video is used to prompt the playback status of the second video during the shooting and recording process. That is, the target video may further include a third video obtained based on the second video synchronously played by a different user device during the shooting and recording process, where the third video displays the playback status of the second video during the shooting and recording process and can participate in subsequent video editing together with the first video.

[0031] For example, multiple users may synchronously watch a second video, i.e., each user device may synchronously play the second video while simultaneously recording each user's viewing experience using the front camera of each user device. Here, the second video may be determined by the initiator of the recording task (also referred to as a master creator). For example, the second video may be obtained based on a second video file or a second video link uploaded by the master creator, and the second video may be synchronously played on all user devices. The playback period and playback state of the second video may be adjusted and controlled, for example, by speeding up the playback speed, slowing down the video, or replaying the video, but this is not limited thereto. The finally obtained third video may be the second video alone, and playback control information for the second video may be further associated with the second video. When the third video is subsequently played, the playback state of the second video may be reproduced based on the playback control information. The third video may also be a recorded version of the second video. For example, the entire playback process of the second video may be recorded to obtain the third video. Specific configurations may be flexible as needed, and are not limited thereto. Then, a third video obtained based on the second video viewed by multiple users is combined with the first video obtained by recording each user's viewing expression, thereby generating a video collaboration in which several people watch the same video content.

[0032] In practice, a target user device can be set, and different user devices can each capture and record a first video for the same recording task. Other user devices can then send the first video they recorded to the target user device, obtain a third video from the target user device, and integrate the first and third videos into the target user device for further processing. Furthermore, each user device can upload the first video they recorded and the third video recorded by the target user device to a server side for further processing, but this is not limited thereto. In some specific embodiments, the different user devices can include a first user device and a second user device, and the first user device can be the device that introduces the second video for the recording task and also the device that initiates the recording task. That is, the first user device can not only initiate the recording task but also introduce the second video for the recording task. Specific configuration can be flexibly configured as needed. If the recording task is not closed and the first user device has not finished the recording task, the target user device may be the first user device. If the recording task is not closed and the first user device has finished the recording task, the target user device may be the second user device. Here, the second user device may be a device other than the first user device among the multiple user devices performing the recording task. It may be manually designated, randomly determined, or determined according to a preset method. For example, the preset method may set the user device invited by the first user device as the second user device. This method effectively ensures the smooth execution of the recording task and video editing process.

[0033] Step S104: Obtain editing templates corresponding to the multiple target videos, where the editing templates include display position information of the multiple target videos and track information of the multiple target videos, where the track information is information about the editing tracks of the multiple target videos in the multi-track editor, such as the type of editing track (main track, picture-in-picture track), etc.

[0034] In practice, multiple editing templates are preset, and a user can select a desired editing template from the multiple preset editing templates and directly use the selected editing template as the editing template corresponding to multiple target videos, or the user (e.g., the initiator of the recording task) can further adjust and modify the selected editing template as needed to obtain editing templates corresponding to multiple target videos. In the embodiments of the present disclosure, the method for obtaining an editing template directly determines the display position and corresponding editing track of each target video, and then the multiple target videos can be easily and quickly edited.

[0035] For example, if the target video includes a third video, the video editing track corresponding to the third video is a main track, and the video editing track corresponding to the first video is a picture-in-picture track. Here, the picture-in-picture track is a secondary track. If multiple target videos include only the first video, the video editing track corresponding to the first video captured by a third user device among different user devices is a main track, and the video editing track corresponding to the first video captured by a user device other than the third user device among the multiple target videos is a picture-in-picture track, where the third user device is the device that initiates the recording task. In practice, the track correspondence method for each target video may be flexibly set, for example, by the user specifying the track corresponding to each video, and is not limited thereto. In practice, the third user device can not only initiate the recording task but also upload the second video. Here, the third user device may be the same user device as the first user device that introduces the second video for the recording task, for example, the master creator's device; furthermore, the third user device may be different from the first user device, for example, the first user device and the third user device are devices of two different users, and the two users can each upload different second videos as needed, but this is not limited thereto.

[0036] Step S106: Display a video editing interface based on the multiple target videos and the editing template.

[0037] In some embodiments, the video editing interface includes a preview playback area and an editing track area. Here, the editing track area includes multiple video editing tracks, including one main track and one or more picture-in-picture tracks (secondary tracks). The multiple video editing tracks are arranged from top to bottom, with the main track located at the top or bottom, but this is not limited thereto. Each target video in the multiple target videos forms a video track clip on a video editing track in the multiple video editing tracks based on the track information, and the timeline positions of the video track clips of the multiple target videos partially or completely overlap. As can be understood, the track information indicates the track type and track position or track indicator corresponding to each target video. Based on the track information of each target video, the video editing track corresponding to each target video in the video editing track is directly determined, and the target video is used as a video track clip on the video editing track to edit the video track clips configured for the multiple target videos. The timeline positions of multiple target videos may overlap partially or completely. For example, if it is assumed that the recording start and end times of all users are the same, the timelines may be considered to completely overlap. Also, if it is assumed that the recording start and end times of some users are different from those of other users, for example, if recording ends early, the resulting target video may partially overlap with the timeline positions corresponding to the target videos of other users. Specific settings may be flexibly set as needed, and are not limited here.

[0038] The images of the multiple target videos are respectively displayed in display areas indicated by the display position information of the multiple target videos in the preview playback area, and the display area of ​​each target video is determined based on the display position information of the target video. With the above method, the display areas of the image screens of the multiple target videos in the preview playback area are directly determined based on the display position information in the editing template, so that the user can clearly understand the presentation effect of each target video through the preview playback area.

[0039] Furthermore, the video editing interface further includes an editing tool operation control for triggering a video editing process in response to a user operation. Here, the editing tool operation control may include, for example, a text editing control, a sticker editing control, an animation special effect control, etc., and may further include a track editing control, but is not limited thereto. In the embodiment of the present disclosure, the arrangement of the above-mentioned several areas of the video editing interface is not limited. For example, the preview playback area may be set above the editing track area, and the operation control of the specified editing tool may be set below the editing track area.

[0040] The above method allows for easy and fast editing of videos shot and recorded by multiple users for the same recording task, effectively reducing the cost of online video collaboration. The device that executes the above video editing method may be a user device or a server, and may interact with other devices during the execution process, and is not limited thereto.

[0041] In some specific embodiments, before the step of obtaining a plurality of recorded videos, the video editing method provided by the embodiments of the present disclosure further includes the following steps (1) to (3).

[0042] (1) In response to receiving a collaborative shooting request from the target user device, create a recording task and display a user interface corresponding to the recording task on the target user device. Here, a video addition control and a user invitation control are displayed on the user interface. Exemplarily, the recording task may be a task for several people to collaboratively watch a video, simultaneously recording an expression video of each user watching the video and a playback of a video (the second video) simultaneously watched by multiple users. In the embodiments of the present disclosure, the specific embodiment of creating the recording task is not particularly limited. For example, an online chat room may be created, and multiple users may enter the online chat room to watch online interactions and / or videos played in the chat room, and simultaneously record each user's online interaction expressions and video playback status.

[0043] (2) In response to the triggering of the video addition control, obtain video information uploaded by the target user device, and obtain a second video based on the video information. Here, the video information includes a local video file and / or a network video link. The second video is a video that multiple users will watch together. In practice, when the target video includes the second video, the target user device may be a device that starts the recording task and simultaneously introduces the second video into the recording task. In some specific embodiments, the target user device may be the same as the first user device, and the first user device can not only start the recording task but also introduce the second video into the recording task.

[0044] In practice, the user may be provided with multiple types of video add controls, such as a video add control for uploading a local video file and / or a video add control for entering a network video link. The user can select the desired video add control as needed. The number of secondary videos may be one or more. For example, the user may upload multiple secondary video files from a local storage device and / or enter multiple network links for secondary videos. In practice, the video information uploaded by the user may be analyzed. For example, when the user uploads a network video link, the network video link may be analyzed to obtain the corresponding secondary video. In practice, if the analysis fails, a prompt may be sent to the user to correct the network video link. Furthermore, to ensure information security, the video information may be verified to prevent the secondary video from containing illegal content or the source of the secondary video from not meeting preset source requirements.

[0045] (3) In response to detecting that the user invitation control has been triggered, generating invitation information for the recording task, and the target user device sending the invitation information to the designated user device, the invitation information being used to invite the designated user device to participate in the recording task.

[0046] Specifically, the user of the target user device (the user who initiates the recording task, also called the master creator) interacts with the designated user by triggering a user invitation control. For example, when the master creator triggers the user invitation control, invitation information such as an invitation passphrase or a chat room link is displayed on the user interface. The master creator forwards the invitation information to other users, and the other users can directly enter the chat room created by the master creator based on the invitation information to perform the recording task.

[0047] For ease of understanding, refer to the schematic diagram of the user interface shown in FIG. 2. After the master creator initiates a collaborative recording request through the target user device, a user interface corresponding to the recording task is displayed on the target user device. The user can upload a video file by triggering the local control, add a video link by triggering the link control, or directly upload a video file by video drag and drop. Furthermore, a preset example video is provided for the user, and the user can directly use and experience the example video as a second video. Furthermore, FIG. 2 also shows the front camera recording window of the master creator (User A in FIG. 2), which is used to display a real-time recording screen including the master creator's face. A user invitation control is located on one side of the front camera recording window of the master creator. The master creator can trigger the user invitation control to obtain invitation information and invite designated users to join the chat room based on the invitation information. At this time, the front camera recording window of the user who was invited to join the chat room is displayed, and the position of the user invitation control is adjusted to the outside of the front camera recording window of the most recently invited user, and the user invitation control is not displayed until the number of users invited by the master creator reaches a preset number threshold. Note that if the master creator does not add videos or invite users, the recording start control on the user interface is in a non-triggerable state, for example, the control is grayed out and does not respond to the user's trigger operation.

[0048] In practice, the user interface of the target user device and the other user devices performing the recording task both display a first area and a second area. Here, the first area is used to display the image screen of the first video captured and recorded by each user device, and the second area is used to display the image screen of the second video. That is, each user participating in the recording task can simultaneously view the second video and can also simultaneously view the user expression screens of all users' front camera recordings.

[0049] For example, referring to the schematic diagram of the user interface shown in FIG. 3, the user interface of the target user device displays front camera recording windows for multiple users (user A to user D), through which image screens of each user's reaction video (i.e., first video) are displayed. Furthermore, the user interface provided by FIG. 3 also displays a second video display window for displaying an image screen of the second video. During this period, the master creator (assumed to be user A) has playback control authority for the second video and can adjust the playback status, such as the progress or speed, of the second video as needed. Since FIG. 3 shows the master creator's video, a recording start control and a second video playback adjustment control are also displayed for the master creator. For participants such as users B to D, their user interface is basically the same as the master creator's user interface, and they can similarly view image screens of all users' reaction videos and second videos. However, the main difference is that the recording start control and the second video playback adjustment control are no longer displayed, i.e., only the master creator has control authority over the recording task. As can be seen from the above user interface, each user in the chat room can simultaneously watch the second video and view each other's reactions, thereby enabling online interaction and even collaborative video shooting by several people.

[0050] As described above, a start recording control is displayed on the user interface of the target user device, and the user (master creator) of the target user device triggers recording through the start recording control. When it is not detected that the target user device triggers the add video control or the user invitation control, the start recording control is in a trigger-disabled state. Here, if the start recording control in a trigger-disabled state is not triggered to record a video, that is, if the master creator does not upload a video or invite other users to participate, it is considered that the recording conditions are not met and recording cannot be started. Therefore, the start recording control is set to a trigger-disabled state, and the user cannot start the recording program through the start recording control.

[0051] When it is detected that the target user device has triggered the add video control and / or the target user device has triggered the user invitation control, the start recording control is in a triggerable state. Here, the start recording control in a triggerable state is triggered to start video recording. That is, when the master creator meets the recording conditions, the master creator can send a recording command through the start recording control to start the recording program. In practice, the start recording control is only displayed on the user interface of the target user device, and not on the user interfaces of other user devices, that is, only the master creator is given the right to control the start recording.

[0052] Based on the above, when obtaining multiple recorded videos, the following steps a and b can be referred to and executed.

[0053] Step a: In response to detecting that the triggerable start recording control is triggered, record and obtain a first video based on frame images collected by the front camera of the user device that executes the recording task. Specifically, after the start recording control is triggered, each user device that executes the recording task starts to record a video of each user's expression, for example, collects a video frame screen including each user's face using the front camera, thereby obtaining a first video corresponding to each user device.

[0054] Step b: When the second video is acquired by the target user device, the user device executing the recording task synchronously plays the second video during the video recording process, records the playback process of the second video, and acquires the third video. For example, while each user device synchronously plays the second video, it uses the front camera to record the user's expression of watching the second video, and the target user device or the server side records the entire playback process of the second video, for example, recording the content area of ​​the second video, and recording the pause, playback, progress bar drag, audio track, etc. of the second video during playback, and acquires the third video.

[0055] In practice, the different user devices include a target user device, and a triggerable video operation control is displayed on a user interface provided by the target user device for the recording task, the video operation control also being referred to as a second video operation control, used to adjust the playback state of the second video; no triggerable video operation control is displayed on a user interface provided by any other user device than the target user device for the recording task; when the video operation control is in the triggerable state, the target user device adjusts the playback state of the second video according to the user's trigger operation in response to a user's trigger operation on the video operation control. That is, the target user device that initiates the recording task is given the operation authority for the second video, and the user of the target user device has the authority to adjust the playback state of the second video during playback, the playback state including, but not limited to, a pause state, a replay state, a fast-forward state, a slow-down state, etc., and the user of the target user device can flexibly control the playback state of the second video as needed, for example, to slow down the playback speed of the important content of the second video or rewatch the important content multiple times when watching the important content of the second video.

[0056] In a specific embodiment, the different user devices include a first user device and a second user device, where the first user device is a device that introduces a second video into the recording task, and when the recording task is not closed and the first user device has not finished the recording task, the target user device is the first user device, and when the recording task is not closed and the first user device has finished the recording task, the target user device is the second user device. In practice, the master creator further has the authority to finish the recording, and based on this, a recording end control is further displayed on the user interface of the first user device, and when the method detects that the first user device has stopped executing the recording task if the recording end control of the first user device is not detected to be triggered, it indicates that the master creator needs to go offline unintentionally or go offline prematurely. Specifically, when the recording task is not closed and the first user device has not yet finished the recording task, a different user device may identify a candidate user device (i.e., the second user device) other than the target user device, and display user-triggered video operation controls and recording termination controls on the user interface displayed when the recording task is executed on the candidate user device. For example, the recording control authority may be handed over to a participant designated by the master creator, or to the first-ranked participant. Specifically, this can be flexibly configured as needed and is not limited thereto. That is, when the master creator goes offline early, the above method can identify a candidate user as a new master creator and perform subsequent video playback control and recording termination operations, thereby fully ensuring the normal execution of the recording task.

[0057] Furthermore, as described above, when multiple target videos are acquired, the embodiments of the present disclosure further provide an embodiment of acquiring editing templates corresponding to multiple target videos, which can be implemented by referring to the following steps A and B.

[0058] Step A: In response to receiving an editing request from a fourth user device on a different user device, display a template selection page. For example, the fourth user device may be the user device that initiates the recording task, the device that uploads the second video, or a specific user device, but is not limited thereto. That is, the fourth user device may be the same as the first user device and / or the third user device. Furthermore, the fourth user device may be another designated device among multiple user devices, but is not limited thereto. In other words, in the embodiment of the present disclosure, the device that initiates the recording task, introduces the second video, and performs video editing may be the same device or a different device, and specific settings may be flexibly configured as needed. Here, the template selection page includes multiple video layout effect diagrams, each of which includes multiple regions, each of which corresponds to a video display position and a video editing track. That is, the video layout effect diagram displays a window display area for each target video, indicating the relative positional relationship between the target videos, and the display positions of the target videos in different video layout effect diagrams are different. In the embodiment of the present disclosure, a video editing track corresponding to each region is set in advance, and the user can automatically determine the display position and video editing track corresponding to each target video by selecting a video layout effect diagram without intentionally setting a video track, which contributes to further improving the efficiency of video editing.

[0059] Step B: In response to a selection operation of a target effect diagram in the plurality of video layout effect diagrams, an editing template corresponding to the target effect diagram is determined as an editing template corresponding to the plurality of target videos.

[0060] The video layout effect chart can intuitively display the display position of each video on the interface and the relative positional relationships between different videos, so that the user can easily and quickly select a desired target effect chart from multiple video layout effect charts. Each video layout effect chart corresponds to an editing template, and after the user selects a target effect chart, the editing template corresponding to the target effect chart can be directly used as the editing template corresponding to multiple target videos. The user can further personally adjust the editing template of the target effect chart as needed and use the adjusted template as the editing template corresponding to multiple target videos.

[0061] In some embodiments, this can be performed with reference to the following steps B1 to B2.

[0062] Step B1: Display layout preview images of multiple target videos according to a target effect diagram selected by a target user from multiple video layout effect diagrams and multiple target videos.

[0063] For ease of understanding, refer to the schematic diagram of multiple layout effects shown in Figure 4. Specifically, six video layout effects are briefly displayed, each indicating the display area and relative positional relationship of a different target video. A user can select a desired target effect diagram as needed, and easily and quickly generate layout preview images of the multiple target videos based on the video display positions and multiple target videos indicated by the target effect diagram, allowing the user to view the display effects of the multiple target videos in subsequent collaborative editing. Furthermore, a user may adjust the proportional size of each video display area indicated by each video layout effect as needed without changing the relative positional relationship of the video display areas, or the user may further adjust and personally set the relative positional relationship of the video display areas based on the selected target effect diagram, but this is not limited thereto.

[0064] Step B2: Determine editing templates corresponding to the plurality of target videos based on the target user's adjustment operation on the display position of the target video in the layout preview image.

[0065] In a specific embodiment, each video layout effect map corresponds to an initial editing template. A layout preview image of multiple target videos is first generated based on the target effect map selected by the target user, the initial editing template corresponding to the target effect map, and multiple target videos. The display position of each target video in the layout preview image is determined based on the display position information of each target video indicated by the initial editing template corresponding to the target effect map. The display position information of the initial editing template corresponding to the target effect map is adjusted based on the target user's adjustment operation for the display position of the target video in the layout preview image to obtain a target editing template, which is an editing template corresponding to the multiple target videos. The main difference between the target editing template and the initial editing template corresponding to the target effect map is the display position of the target video adjusted by the user.

[0066] An embodiment of the present disclosure further provides an operation method for a user to adjust the display position of a target video. For example, the user can adjust the display position of a target video by dragging and dropping the floating window of the target video. Furthermore, a horizontal layout and a vertical layout are preset. When one or more floating windows of the target video are dragged into a first designated area, a horizontal layout is displayed to the user. For example, a plurality of first videos are horizontally arranged above a third video. When one or more floating windows are dragged into a second designated area, a vertical layout is displayed to the user. For example, a plurality of first videos are vertically arranged to the sides of the third video. The above method allows the layout of multiple target videos to be easily and quickly adjusted and editing templates corresponding to multiple target videos to be quickly determined.

[0067] After determining editing templates (i.e., the target editing templates) corresponding to the multiple target videos, the target videos are introduced into the target editing templates to generate a video editing draft. The video editing draft is regarded as a process file and includes materials, such as videos, and editing operation information for the materials. The video editing draft is then introduced into a multi-track editor. In other words, the editing draft displayed on the multi-track editor is obtained by introducing the target videos into the target editing templates. A video editing interface is displayed to the user, and the user directly edits the multiple target videos through the video editing interface. Finally, an edited fourth video is generated, which is also known as video co-shooting or video co-creation.

[0068] The above method significantly reduces the editing costs of online video collaboration, and multiple users, regardless of geographical location constraints, can easily and quickly achieve the goal of video collaboration by simply using one device and launching a software application capable of executing the video editing method provided by the embodiment of the present disclosure. For ease of understanding, the embodiment of the present disclosure further provides a flowchart of the video editing method shown in FIG. 5, which illustrates the operation flows before, during, and after recording by the master creator (corresponding to the target user device that initiates the recording task) and the participants (corresponding to the invited user devices). Here, before recording, the master creator triggers the creation of a chat room through the collaborative recording function entry, introduces all videos (the second videos) that users will jointly watch when performing the recording task on the chat room interface, and then interacts with participants by triggering user invitation controls, etc. After receiving the invitation, the participants directly enter the chat room based on the invitation information. During the recording process, the master creator has the authority to start recording, play / pause the co-viewing video, and confirm when the recording is finished. Participants simply follow the master creator to watch the video and then complete the recording. After recording, the master creator can select a template and download participants' reaction videos (i.e., the participants' expressive videos recorded by their devices as they watch the video) or directly upload their own reaction videos. The master creator can download each participant's video in the background, import all the recorded videos into a multi-track editor for editing, and finally post the final edited video. This method makes online video co-creation easy and fast, significantly reducing the cost of multi-person video co-creation and effectively improving the user experience.

[0069] Corresponding to the video editing method, FIG. 6 is a structural schematic diagram of a video editing device provided by an embodiment of the present disclosure. The device can be realized by software and / or hardware and generally integrated into electronic devices. As shown in FIG. 6, the video editing device includes: a video acquisition module 602 for acquiring a plurality of target videos, the plurality of target videos including first videos respectively captured and recorded by different user devices for the same recording task; a template acquisition module 604 for acquiring an editing template corresponding to a plurality of target videos, the editing template including display position information of the plurality of target videos and track information of the plurality of target videos; a video editing module 606 for displaying a video editing interface based on the plurality of target videos and the editing template, wherein: The video editing interface includes a preview playback area and an edit track area, the edit track area includes a plurality of video edit tracks, images of the plurality of target videos are respectively displayed in display areas indicated by respective display position information of the plurality of target videos in the preview playback area, each target video in the plurality of target videos forms a video track clip on a video edit track in the plurality of video edit tracks respectively based on track information, and the timeline positions of the video track clips of the plurality of target videos partially or fully overlap.

[0070] The above device allows multiple users to easily and quickly edit videos shot and recorded in a centralized manner, effectively reducing the cost of online video collaboration.

[0071] In some embodiments, the target video further includes a third video obtained based on a second video, the second video being a video played back during the process of being filmed and recorded by the different user device for the recording task, and the third video being used to represent the playback state of the second video during the filming and recording process.

[0072] In some embodiments, the different user device includes a target user device, and a video operation control in a triggerable state is displayed on a user interface provided by the target user device for the recording task, and the video operation control in a triggerable state is not displayed on a user interface provided by any other user device other than the target user device among the different user devices for the recording task, and when the video operation control is in a triggerable state, in response to detecting a user trigger operation on the video operation control on the target user device, the playback state of the second video is adjusted based on the user trigger operation.

[0073] In some embodiments, the different user devices include a first user device and a second user device, the first user device being a device that introduces the second video into the recording task, and when the recording task is not closed and the first user device has not ended the recording task, the target user device is the first user device, and when the recording task is not closed and the first user device has ended the recording task, the target user device is the second user device.

[0074] In some embodiments, the video editing track corresponding to the tertiary video is a main track and the video editing track corresponding to the primary video is a picture-in-picture track.

[0075] In some embodiments, when the plurality of target videos includes only a first video, the video editing track corresponding to the first video shot by a third user device among the different user devices is a main track, and the video editing track corresponding to the first video shot by a user device other than the third user device in the plurality of target videos is a picture-in-picture track, wherein the third user device is the device that initiates the recording task.

[0076] In some embodiments, the template acquisition module 604 is specifically used to display a template selection page in response to receiving an editing request from a fourth user device on the different user device, where the template selection page includes a plurality of video layout effect diagrams, the video layout effect diagrams including a plurality of regions, each region in the plurality of regions corresponding to a video display position and a video editing track, and in response to detecting a selection operation on a target effect diagram in the plurality of video layout effect diagrams, determine an editing template corresponding to the target effect diagram as an editing template corresponding to the plurality of target videos.

[0077] In some embodiments, the template acquisition module 604 is specifically used to display layout preview images of multiple target videos based on a target effect diagram and multiple target videos selected by a target user from multiple video layout effect diagrams, and to determine editing templates corresponding to the multiple target videos based on the target user's adjustment operation on the display position of the target video in the layout preview image.

[0078] In some embodiments, the device comprises: a task creation module for creating a recording task and displaying a user interface corresponding to the recording task on the target user device in response to receiving the collaborative recording request of the target user device, wherein an add video control and a user invitation control are displayed on the user interface; a second video acquisition module for acquiring video information uploaded by the target user device in response to detecting that the video add control has been triggered, and acquiring a second video based on the video information, where the video information includes a local video file and / or a network video link; The system further includes a task invitation module for generating invitation information for a recording task in response to detecting that a user invitation control has been triggered, and for transmitting the invitation information to a user device designated by the target user device, wherein the invitation information is used to invite the designated user device to participate in the recording task.

[0079] In some embodiments, a first area and a second area are displayed on the user interface of both the target user device and the other user device performing the recording task, where the first area is used to display an image screen of a first video captured and recorded by each user device, and the second area is used to display an image screen of a second video.

[0080] In some embodiments, a start recording control is further displayed on the user interface of the target user device, and when it is not detected that the target user device triggers the add video control or the target user device triggers the user invitation control, the start recording control is in a non-triggerable state, wherein the start recording control in the non-triggerable state is not triggered to record video, and when it is detected that the target user device triggers the add video control and / or the target user device triggers the user invitation control, the start recording control is in a triggerable state, wherein the start recording control in the triggerable state is triggered to start video recording.

[0081] In some embodiments, the video acquisition module 602 is specifically used to record and acquire a first video based on frame images collected by the front camera of the user device performing the recording task in response to the triggering of a recording start control in a triggerable state, and when a second video is acquired by the target user device, to synchronously play back the second video during the video recording process through the user device performing the recording task, record the playback process of the second video, and acquire a third video.

[0082] The video editing device provided by the embodiments of the present disclosure implements the video editing method provided by any of the embodiments of the present disclosure, and has corresponding function modules and beneficial effects for implementing the method.

[0083] For the sake of simplicity and brevity, those skilled in the art can refer to the corresponding steps in the method embodiments for the specific work steps of the above-described apparatus embodiments, and the detailed steps will not be repeated here.

[0084] An embodiment of the present disclosure further provides an electronic device, the electronic device comprising: a processor; and a memory for storing processor-executable instructions, the processor being used to implement the video editing method by reading the executable instructions from the memory and executing the instructions.

[0085] 7 is a structural schematic diagram of an electronic device provided by an embodiment of the present disclosure. As shown in FIG. 7, an electronic device 700 includes one or more processors 701 and a memory 702.

[0086] The processor 701 may be a central processing unit (CPU) or other type of processing unit having data processing and / or instruction execution capabilities, and may also control other components in the electronic device 700 to perform desired functions.

[0087] The memory 702 may include one or more computer program products. The computer program products may be in various forms of computer-readable storage media, such as volatile memory and / or non-volatile memory. The volatile memory may include, for example, random access memory (RAM) and / or cache memory. The non-volatile memory may include, for example, read-only memory (ROM), a hard disk, or flash memory. One or more computer program instructions are stored in the computer-readable storage medium, and the processor 701 executes the program instructions to implement the video editing method of the embodiments of the present disclosure described above and / or other desired functions. The computer-readable storage medium may also store various contents, such as an input signal, signal components, and noise components.

[0088] In one example, the electronic device 700 further comprises an input device 703 and an output device 704, which are connected to each other via a bus system and / or other type of connection mechanism (not shown).

[0089] Furthermore, the input device 703 may include, for example, a keyboard, a mouse, and the like.

[0090] The output device 704 can output various information such as determined distance information, direction information, etc. The output device 704 may include, for example, a display, a speaker, a printer, a communication network, and a remote output device connected thereto.

[0091] Of course, for simplicity, Fig. 7 shows only some components of the electronic device 700 that are relevant to the present disclosure, and omits components such as buses, input / output interfaces, etc. In addition, depending on a specific application, the electronic device 700 may further include any other appropriate components.

[0092] In addition to the above methods and devices, embodiments of the present disclosure may also be a computer program product that includes computer program instructions that, when executed by a processor, cause the processor to perform the video editing methods provided by embodiments of the present disclosure.

[0093] The computer program product may have program code written in any combination of one or more program design languages ​​for carrying out operations of embodiments of the present disclosure, including object-oriented program design languages ​​such as Java, C++, etc., and also including conventional procedural program design languages ​​such as "C" or similar program design languages. The program code may execute entirely on the user computing device, partially on the user device, as a separate software package, partially on the user computing device and partially on a remote computing device, or entirely on a remote computing device or server.

[0094] Furthermore, an embodiment of the present disclosure may be a computer-readable storage medium having stored thereon computer program instructions that, when executed by a processor, cause the processor to perform a video editing method provided by an embodiment of the present disclosure.

[0095] The computer-readable storage medium may be any combination of one or more computer-readable media. The computer-readable medium may be a readable signal medium or a readable storage medium. The computer-readable storage medium may include, but is not limited to, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any combination thereof. More specific examples (non-exhaustive list) of computer-readable storage media include, but are not limited to, an electrical connection having one or more wires, a portable disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the above.

[0096] An embodiment of the present disclosure further provides a computer program product, which includes computer programs / instructions, which, when executed by a processor, implement the video editing method in an embodiment of the present disclosure.

[0097] It should be noted that, in this specification, relational terms such as "first" and "second" are used to distinguish one entity or operation from another, and do not require or imply the existence of any actual relationship or ordering between those entities or operations. Furthermore, the terms "comprise," "include," or any other variation thereof cover non-exclusive inclusions, such that a process, method, article, or device comprising a set of elements includes, in addition to those elements, other elements not expressly listed or inherent in the process, method, article, or device. Unless more specific, an element defined by the phrase "comprising..." does not exclude the presence of other identical elements in the process, method, article, or device that comprises said element.

[0098] Specific embodiments of the present disclosure have been described above so that those skilled in the art can fully understand or realize the present disclosure. Many modifications of these examples will be apparent to those skilled in the art, and the general principles defined herein may be implemented in other examples without departing from the spirit or scope of the present disclosure. Therefore, the present disclosure is not intended to be limited to the examples described herein, but is to be accorded the widest scope consistent with the principles and novel features disclosed herein.

Claims

1. 1. A video editing method comprising: acquiring a plurality of target videos; the plurality of target videos include first videos captured and recorded by different user devices for the same recording task; obtaining editing templates corresponding to the plurality of target videos; the editing template includes display position information of the plurality of target videos and track information of the plurality of target videos; The step of obtaining editing templates corresponding to the plurality of target videos includes: displaying a template selection page including a plurality of video layout effect diagrams in response to an editing request from one of the different user devices; each video layout effect diagram includes a plurality of regions, and each of the plurality of regions corresponds to a video display position and a video editing track; responsive to detecting a selection of a target effect diagram from the plurality of video layout effect diagrams, determining editing templates corresponding to the plurality of target videos based on the editing templates corresponding to the target effect diagrams; Including, displaying a video editing interface based on the plurality of target videos and the editing template; the video editing interface includes a preview playback area and an edit track area; the edit track area includes a plurality of video edit tracks; The images of the plurality of target videos are respectively displayed in the preview playback area indicated by the display position information of each of the plurality of target videos; each of the plurality of target videos forms a video track segment in one of the plurality of video editing tracks based on the track information; and the timeline positions of the video track segments of the plurality of target videos partially or completely overlap; Equipped with method.

2. the target video further includes a third video acquired based on the second video; the second video is a video played by the different user device during a shooting and recording process for the recording task; The third video is configured to display the playback status of the second video during the shooting and recording process. The method of claim 1.

3. the different user device includes a target user device; a user interface provided by the target user device for the recording task is displayed with triggerable video manipulation controls; user interfaces provided by other user devices among the different user devices, except for the target user device for the recording task, are not displayed with the video manipulation control in the triggerable state; and when the video operation control is in the triggerable state, in response to detecting a user trigger operation on the video operation control at the target user device, adjusting a playback state of the second video based on the user trigger operation. The method of claim 2.

4. the different user devices include a first user device and a second user device; the first user device is a device that introduces the second video for the recording task; If the recording task is not closed and the first user device has not finished the recording task, the target user device is the first user device; If the recording task is not closed and the first user device ends the recording task, the target user device is the second user device. The method of claim 3.

5. the video editing track corresponding to the third video is a main track, and the video editing track corresponding to the first video is a picture-in-picture track; The method of claim 2.

6. The method further comprises: Before you get multiple recording videos, In response to receiving a collaborative shooting request from a target user device, creating the recording task and displaying a user interface corresponding to the recording task on the target user device; the user interface displays an add video control and a user invitation control; In response to the video add control being triggered, acquiring video information uploaded by the target user device, and acquiring a second video based on the video information; the video information includes a local video file and / or a network video link; generating invitation information for the recording task for the target user device in response to the user invitation control being triggered, and transmitting the invitation information to the designated user device; the invitation information is configured to invite the designated user device to participate in the recording task; Including, The method of claim 2.

7. a first area and a second area are displayed on the user interface of the target user device and the other user device that executes the recording task; The first area is configured to display an image screen of the first video captured and recorded by each user device; and the second area is configured to display an image screen of the second video; The method of claim 6.

8. the user interface of the target user device further displays a start recording control; When the add video control is not detected as being triggered by the target user device and the user invitation control is not detected as being triggered by the target user device, the start recording control in a non-triggerable state cannot be triggered to record a video; When the add video control is detected as being triggered by the target user device and / or the user invitation control is detected as being triggered by the target user device, the start recording control is in a triggerable state, and the start recording control in a triggerable state is configured to start recording a video when triggered. The method of claim 6.

9. The step of obtaining the plurality of recorded videos includes: In response to detecting that a recording start control in a triggerable state has been triggered, performing recording based on frame images collected by a front camera of a user device performing the recording task to obtain a first video; When a second video is acquired by the target user device, playing the second video and recording the playing process of the second video by the user device that executes the recording task during the video recording process to acquire a third video; The method of claim 8, comprising:

10. When the plurality of target videos only includes the first video, a video editing track corresponding to the first video captured by a third user device among the different user devices is a main track; and a video editing track corresponding to the first video in the plurality of target videos, the video editing track being shot by a user device other than the third user device, is a picture-in-picture track; the third user device is a device that initiates the recording task; The method of claim 1.

11. determining the editing templates corresponding to the plurality of target videos based on the editing templates corresponding to the target effect diagrams, displaying a layout preview image of the plurality of target videos according to the target effect diagram and the plurality of target videos; determining an editing template corresponding to the plurality of target videos based on an operation of adjusting the display position of the target video in the layout preview image; The method of claim 1 , comprising:

12. 1. An electronic device, comprising: the electronic device includes a processor and a memory for storing processor-executable instructions; The processor is configured to read the instructions from the memory and execute the instructions to perform operations, the operations including: acquiring a plurality of target videos; the plurality of target videos include first videos captured and recorded by different user devices for the same recording task; obtaining editing templates corresponding to the plurality of target videos; the editing template includes display position information of the plurality of target videos and track information of the plurality of target videos; The step of obtaining editing templates corresponding to the plurality of target videos includes: displaying a template selection page including a plurality of video layout effect diagrams in response to an editing request from one of the different user devices; each video layout effect diagram includes a plurality of regions, and each of the plurality of regions corresponds to a video display position and a video editing track; responsive to detecting a selection of a target effect diagram from the plurality of video layout effect diagrams, determining editing templates corresponding to the plurality of target videos based on the editing templates corresponding to the target effect diagrams; Including, displaying a video editing interface based on the plurality of target videos and the editing template; the video editing interface includes a preview playback area and an edit track area; the edit track area includes a plurality of video edit tracks; The images of the plurality of target videos are respectively displayed in the preview playback area indicated by the display position information of each of the plurality of target videos; each of the plurality of target videos forms a video track segment in one of the plurality of video editing tracks based on the track information; and the timeline positions of the video track segments of the plurality of target videos partially or completely overlap; Equipped with electronic equipment.

13. the target video further includes a third video acquired based on the second video; the second video is a video played by the different user device during a shooting and recording process for the recording task; The third video is configured to display the playback status of the second video during the shooting and recording process.

13. The electronic device of claim 12.

14. When the plurality of target videos only includes the first video, a video editing track corresponding to the first video captured by a third user device among the different user devices is a main track; and a video editing track corresponding to the first video in the plurality of target videos, the video editing track being shot by a user device other than the third user device, is a picture-in-picture track; the third user device is a device that initiates the recording task; 13. The electronic device of claim 12.

15. A non-transitory computer-readable storage medium, comprising: The storage medium stores a computer program; The computer program, when executed by a processor, causes the processor to perform operations, the operations including: acquiring a plurality of target videos; the plurality of target videos include first videos captured and recorded by different user devices for the same recording task; obtaining editing templates corresponding to the plurality of target videos; the editing template includes display position information of the plurality of target videos and track information of the plurality of target videos; The step of obtaining editing templates corresponding to the plurality of target videos includes: displaying a template selection page including a plurality of video layout effect diagrams in response to an editing request from one of the different user devices; each video layout effect diagram includes a plurality of regions, and each of the plurality of regions corresponds to a video display position and a video editing track; responsive to detecting a selection of a target effect diagram from the plurality of video layout effect diagrams, determining editing templates corresponding to the plurality of target videos based on the editing templates corresponding to the target effect diagrams; Including, displaying a video editing interface based on the plurality of target videos and the editing template; the video editing interface includes a preview playback area and an edit track area; the edit track area includes a plurality of video edit tracks; The images of the plurality of target videos are respectively displayed in the preview playback area indicated by the display position information of each of the plurality of target videos; each of the plurality of target videos forms a video track segment in one of the plurality of video editing tracks based on the track information; and the timeline positions of the video track segments of the plurality of target videos partially or completely overlap; Equipped with A non-transitory computer-readable storage medium.

16. the target video further includes a third video acquired based on the second video; the second video is a video played by the different user device during a shooting and recording process for the recording task; The third video is configured to display the playback status of the second video during the shooting and recording process.

16. The non-transitory computer-readable storage medium of claim 15.

17. When the plurality of target videos only includes the first video, a video editing track corresponding to the first video captured by a third user device among the different user devices is a main track; and a video editing track corresponding to the first video in the plurality of target videos, the video editing track being shot by a user device other than the third user device, is a picture-in-picture track; the third user device is a device that initiates the recording task; 16. The non-transitory computer-readable storage medium of claim 15.

Citation Information

Patent Citations

  • Collaborative Digital Video Platform That Enables Synchronized Capture, Curation And Editing Of Multiple User-Generated Videos

    US20140186004A1

  • Live group video streaming

    WO2022051083A1