Video processing method and device, storage medium, electronic device and program product

By presenting the identification box associated with member locations on the video playback page, it solves the problem that users find it difficult to accurately select specific members in the video, and achieves efficient and intuitive interactive selection, improving the quality and user experience of video interaction.

CN119729136BActive Publication Date: 2025-08-26BEIJING DAJIA INTERNET INFORMATION TECH CO LTD
View PDF 2 Cites 0 Cited by

Patent Information

Application Number
CN202510238971.8
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2025-03-03
Publication Date
2025-08-26
Estimated Expiration
2045-03-03

AI Technical Summary

Technical Problem

In video interactive scenarios, it is difficult for users to accurately match the avatar and nickname in the interactive interface with the real person in the video, resulting in increased selection difficulty and reducing the convenience and accuracy of interaction.

Method used

The identification box associated with member positions is presented on the video playback page, and the interactive target members are determined through the identification box, and the dimensions and position of the identification box are adjusted dynamically to achieve accurate identification and selection.

Benefits of technology

It improves the efficiency and accuracy of interaction selection, simplifies the operation process, enhances the interactive experience between users and video content, and improves the quality of video interaction and user satisfaction.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN119729136B_ABST
    Figure CN119729136B_ABST
Patent Text Reader

Abstract

The present application provides a video processing method and device, a storage medium, an electronic device and a program product, which relate to the field of computer technology. The video processing method includes: in response to an interactive instruction, presenting identification boxes corresponding to multiple members in the target video on the playback page of the target video, the identification box corresponding to each member is associated with the position of the member in the playback page, and the interactive instruction is used to initiate an interactive request to at least one member in the target video; in response to a selection instruction based on the identification box, determining the target member among the multiple members. The solution in the present application allows users to determine the target member for interaction based on the identification box, without having to rely on static nicknames and avatars for selection, which significantly improves the selection efficiency and accuracy of the interaction.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] The present application relates to the field of computer technology, and in particular to a video processing method and apparatus, a storage medium, an electronic device, and a program product. Background Art

[0002] In existing video interaction scenarios, when a user wants to interact with a specific member in a video that includes multiple members, he or she usually needs to select the member through the nickname and avatar in the interaction interface.

[0003] However, since the positions and postures of the members in the video are constantly changing, it is difficult for users to accurately match the avatars and nicknames in the interactive interface with the real people in the video, which increases the difficulty of user selection and reduces the convenience and accuracy of interaction. Summary of the Invention

[0004] In view of this, embodiments of the present application provide a video processing method and apparatus, a storage medium, an electronic device, and a program product.

[0005] In a first aspect, an embodiment of the present application provides a video processing method, which is applied to a first user terminal, and the method includes: in response to an interactive instruction, presenting identification boxes corresponding to multiple members in the target video on the playback page of the target video, the identification box corresponding to each member is associated with the position of the member in the playback page, and the interactive instruction is used to initiate an interactive request to at least one member in the target video; in response to a selection instruction based on the identification box, determining the target member among the multiple members.

[0006] In combination with the first aspect, in some implementations of the first aspect, the identification frame corresponding to each member is located in the facial area of ​​the member, and / or the identification frame corresponding to each member is located in a specific area above the head of the member.

[0007] In combination with the first aspect, in certain implementations of the first aspect, the identification box corresponding to each member follows the position of the member in the playback page; and / or, the size of the identification box corresponding to each member is positively correlated with the size of the member in the playback page.

[0008] In combination with the first aspect, in certain implementations of the first aspect, the video processing method further includes: identifying facial information of each of the multiple members in the target video; labeling the multiple members based on their facial information, and generating identification boxes for each of the multiple members.

[0009] In combination with the first aspect, in certain implementations of the first aspect, the video processing method further includes: obtaining identification information of each of the multiple members in the target video from the second user terminal; labeling the multiple members based on the identification information of the multiple members, and generating identification boxes for the multiple members.

[0010] In combination with the first aspect, in certain implementations of the first aspect, identification information of each of the multiple members in the second user terminal is generated by the second user terminal before the target video is played based on facial images of the members to be included in the target video.

[0011] In combination with the first aspect, in certain implementations of the first aspect, the video processing method further includes: determining a target special effect, where the target special effect refers to a visual effect and / or audio effect triggered during the playback of the target video; and playing the target special effect based on the target member.

[0012] In combination with the first aspect, in some implementations of the first aspect, playing a target special effect based on a target member includes: playing the target special effect at a designated position corresponding to the target member in a playback page.

[0013] In combination with the first aspect, in certain implementations of the first aspect, a target special effect is played at a designated position corresponding to the target member in the playback page, including: if there is only one target member, the target member is given an enlarged close-up display, and the target special effect is played at a designated position corresponding to the target member in the playback page; if there are multiple target members, a special effect movement trajectory is generated based on the designated positions corresponding to the multiple target members, and the target special effect is played based on the special effect movement trajectory.

[0014] In combination with the first aspect, in certain implementations of the first aspect, the target special effects include a gift-giving special effect, and the target special effects are played at a designated position corresponding to the target member in the playback page, including: adding the target member's identity information to the visual elements of the gift-giving special effect to generate a customized gift special effect containing the target member's identity information; playing a customized gift special effect containing the target member's identity information at a designated position corresponding to the target member in the playback page.

[0015] In combination with the first aspect, in certain implementations of the first aspect, the target special effect includes a gift-giving special effect, and the video processing method further includes: presenting a gift panel including multiple gifts on the playback page of the target video so that the user can determine the gift-giving special effect based on the gift panel.

[0016] In combination with the first aspect, in certain implementations of the first aspect, before determining the target member among multiple members in response to a selection instruction based on the identification box, it also includes: when at least one member is selected based on the identification box, selection prompt information corresponding to at least one member is presented on the playback page of the target video so that the user can determine whether to select the at least one member as the target member. The selection prompt information refers to the prompt information displayed on the playback page of the target video on whether the at least one member is selected.

[0017] In the second aspect, an embodiment of the present application provides a video processing device, which is applied to a first user end, and the device includes: a presentation module, which is used to present, on the playback page of the target video, identification boxes corresponding to multiple members in the target video in response to the user's interaction instructions, and the identification box corresponding to each member is associated with the position of the member in the playback page, and the interaction instruction is used to initiate an interaction request to at least one member in the target video; a determination module, which is used to determine the target member among the multiple members in response to the user's selection instruction based on the identification box.

[0018] In a third aspect, an embodiment of the present application provides a computer-readable storage medium, which stores a computer program for executing the video processing method described in the first aspect.

[0019] In a fourth aspect, an embodiment of the present application provides an electronic device, comprising: a processor; a memory for storing processor-executable instructions; and the processor is configured to execute the video processing method described in the first aspect.

[0020] In a fifth aspect, an embodiment of the present application provides a computer program product, which includes instructions. When the instructions are executed on an electronic device, the electronic device implements the video processing method described in the first aspect.

[0021] In this application, by presenting an identification box associated with the member's location on the target video playback page, users can determine the target member for interaction based on the identification box, eliminating the need to rely on static nicknames and avatars for selection, significantly improving the efficiency and accuracy of interactive selection. In addition, this intuitive selection method simplifies the operation process, enhances the user's interactive experience with video content, enables users to interact with target members more conveniently and accurately, and improves the overall quality of video interaction and user satisfaction. BRIEF DESCRIPTION OF THE DRAWINGS

[0022] The above and other purposes, features, and advantages of the present application will become more apparent through a more detailed description of the embodiments of the present application in conjunction with the accompanying drawings. The accompanying drawings are intended to provide a further understanding of the embodiments of the present application and constitute a part of the specification. Together with the embodiments of the present application, they are used to explain the present application and do not constitute a limitation of the present application. In the drawings, the same reference numerals generally represent the same components or steps.

[0023] Figure 1 Shown is a schematic diagram of a video interactive page in the prior art.

[0024] Figure 2 FIG2 is a schematic diagram of an implementation environment of a video processing method provided in an embodiment of the present application.

[0025] Figure 3FIG2 is a flow chart of a video processing method provided in an embodiment of the present application.

[0026] Figure 4 Shown is a schematic diagram of a playback page of a target video in a live broadcast scenario provided by an embodiment of the present application.

[0027] Figure 5 Shown is a schematic diagram of a playback page of a target video in a multi-person PK scenario provided by an embodiment of the present application.

[0028] Figure 6 Shown is a schematic diagram of a playback target special effect provided by an embodiment of the present application.

[0029] Figure 7 Shown is a schematic diagram of playback target special effects provided by another embodiment of the present application.

[0030] Figure 8 FIG2 is a schematic diagram of the structure of a video processing device provided in one embodiment of the present application.

[0031] Figure 9 Shown is a structural schematic diagram of an electronic device provided in one embodiment of the present application. DETAILED DESCRIPTION

[0032] The following will be combined with the drawings in the embodiments of this application to clearly and completely describe the technical solutions in the embodiments of this application. Obviously, the embodiments described are only part of the embodiments of this application, not all of the embodiments. Based on the embodiments in this application, all other embodiments obtained by ordinary technicians in this field without making creative efforts are within the scope of protection of this application.

[0033] In today's video interaction environment, user engagement and user experience are crucial. However, existing interaction mechanisms still have obvious limitations in some aspects, especially when users want to interact with specific members in the video, such as expressing support, sending virtual gifts, or other forms of personalized communication.

[0034] Figure 1 The figure shows a schematic diagram of a video interactive page in the prior art. Figure 1 As shown in the left figure, when a user wants to send a virtual gift to a member in the video, he usually needs to operate on the dedicated video interaction page. Specifically, when the user expands the folded option corresponding to "Send Gift to", the page will list the following Figure 1 The picture on the right shows the nicknames and avatars of all members participating in the video interaction, and users need to select the one they want to interact with.

[0035] However, this selection process has some flaws. Specifically, participants in the video may appear in different positions and angles, and the footage is dynamic, with participants' positions and postures constantly changing as the video progresses. This makes it difficult for users to accurately match the static avatars and nicknames in the interactive interface with the dynamic real-life images in the video. This ambiguity not only increases user uncertainty when selecting an interaction partner, but can also lead to incorrect selections, impacting interaction accuracy and user satisfaction.

[0036] In summary, existing video interaction mechanisms suffer from limitations in accuracy and intuitiveness when users select specific members. Therefore, this application provides a video processing method that can provide a more intuitive and accurate selection mechanism, thereby improving the overall quality of video interaction. The following details the implementation environment and specific implementation methods of this video processing method.

[0037] Figure 2 FIG. 1 is a schematic diagram of an implementation environment of a video processing method provided by an embodiment of the present application. Figure 2 As shown, the implementation environment includes a first user terminal 210, a second user terminal 220, and a server 230. For example, the first user terminal 210 is a user terminal (also known as a viewer terminal) corresponding to a user who wants to interact with a member of a target video, and the second user terminal 220 is a user terminal (also known as a host terminal) corresponding to multiple members of the target video. The first and second user terminals 210 and 220 are connected to the server 230 via a wireless or wired network.

[0038] In some embodiments, the first user terminal 210 and the second user terminal 220 include smartphones, tablet computers, laptop computers, desktop computers, etc. The server 230 is an independent physical server, or a server cluster or distributed system composed of multiple physical servers, or a cloud server that provides basic cloud computing services such as cloud services, cloud databases, cloud computing, cloud functions, cloud storage, network services, cloud communications, middleware services, domain name services, security services, CDN (Content Delivery Network), and big data and artificial intelligence platforms, etc., which are not limited in this disclosure.

[0039] In an example scenario of a live group broadcast, the screen of the first user terminal 210 displays multiple group broadcast members. When a user wishes to interact with a member, they simply issue an interaction command, and a marker box precisely corresponding to the group broadcast member's position appears in real time on the live broadcast screen. For example, these marker boxes move with the member's movements. By clicking on a marker box, the user can select the target member identified by the marker box. This method makes user selection more intuitive and accurate, greatly improving the gift-giving experience.

[0040] In another interactive scenario during a multi-person video call, a user wants to send a fun virtual sticker to Friend B to express their feelings. After the user issues an interactive command on the video chat interface, multiple markers corresponding to the positions of the friends appear on the screen. The user clicks the marker corresponding to Friend B, selecting Friend B as the interaction partner. Then, they select a virtual sticker and click Send, and the sticker appears on Friend B's face. This process ensures precise alignment between the markers and the friend's position, making user operations more intuitive and accurate, and enhancing the interactivity and fun of video chats.

[0041] Figure 3 FIG. 1 is a flow chart of a video processing method according to an embodiment of the present application. Figure 3 As shown, the video processing method includes the following steps.

[0042] Step S310 , in response to the interactive instruction, presenting identification boxes corresponding to the multiple members in the target video on the playback page of the target video.

[0043] The interactive instruction is used to initiate an interactive request to at least one member in the target video. For example, the interactive instruction refers to a request issued by the user to interact with the member in the target video by operating the playback page while watching the target video, such as clicking, touching, or voice command.

[0044] After receiving the interactive instruction, the first user terminal generates and displays an identification box for each of the multiple members appearing in the target video on the playback page of the target video. Each identification box corresponding to each member is associated with the member's position on the playback page. An identification box is a visual element used to highlight and identify a specific member on the playback page. Optionally, the identification box is a rectangular, circular, or other shaped border that surrounds a designated location corresponding to the target member, such as the head or chest.

[0045] Step S320 : determining a target member among the multiple members in response to a selection instruction based on the identification box.

[0046] It is understandable that the purpose of the identification box is to help users identify and select the members they want to interact with more intuitively and accurately. Therefore, when the user sees the identification box associated with the member's position, he or she can issue a selection instruction by clicking or touching. Then, the first user terminal determines the target member from multiple members based on the user's selection instruction. Optionally, the target member includes members who interact with the user, such as those who send virtual gifts, like or comment. Optionally, the target member also includes members who perform specific actions (such as waving, raising hands, etc.) or trigger certain preset events (such as entering the screen, leaving the screen, performing specific performances, etc.).

[0047] In this embodiment, by displaying an identification box associated with the member's location on the target video's playback page, users can identify the target member for interaction based on the identification box, eliminating the need to rely on static nicknames and profile pictures. This significantly improves the efficiency and accuracy of interactive selection. Furthermore, this intuitive selection method simplifies the operation process, enhances the user's interactive experience with video content, and enables users to interact with target members more conveniently and accurately, improving the overall quality of video interaction and user satisfaction.

[0048] exist Figure 3 Based on the illustrated embodiment, the video processing method further includes: when at least one member is selected based on the identification box, selection prompt information corresponding to at least one member is presented on the playback page of the target video to determine whether the at least one member is selected as the target member.

[0049] Selection prompt information refers to the prompt information displayed on the playback page of the target video, indicating whether the at least one member is selected. Exemplarily, selection prompt information refers to the interactive prompt displayed on the playback page when the user selects at least one member in the target video through the identification box. The purpose is to clearly indicate the current selection status to the user and guide the user to confirm whether to determine the selected member as the target member. Exemplarily, the selection prompt information can be a text prompt, a graphic logo, an animation effect or any other form of visual prompt to enhance the intuitiveness and accuracy of the user's operation and ensure that the user can clearly understand the result of the current operation when selecting the target member.

[0050] Figure 4 The figure shows a schematic diagram of the target video playback page in a live broadcast scenario provided by an embodiment of the present application. Figure 4In the left image, each member's identification box is black and located in the member's header area. When the user selects at least one member based on the identification box, the selected member's identification box turns red. This change from black to red serves as a selection prompt. This visual change provides users with feedback on the current selection status and clearly informs them which members have been selected. This design enhances user intuitiveness and simplifies the selection process, allowing users to quickly and accurately confirm whether the selected member is the target member.

[0051] Figure 5 The figure shows a schematic diagram of the target video playback page in a multiplayer PK (Player Kill) scenario provided by an embodiment of the present application. Figure 5 In the scene shown, four streamers appear simultaneously in the video. Each streamer has a marker above their head, and users can click on the marker to select the streamer they want to interact with. As shown in the image on the right, when a user selects a streamer, the marker turns red, clearly indicating that the streamer has been selected for interaction. If the user wants to add another streamer to the interaction, they simply click on the marker of another streamer, and the marker of that streamer will also turn red, ensuring that the user can clearly see their selection.

[0052] In addition, when the user clicks the identification box, the member is selected as the target member. If the user wants to cancel, just click the identification box of the member again. Figure 4 and Figure 5 In the playback page, for the selected anchors, the avatars and nicknames of these anchors will be displayed, further clarifying the user's selection object and providing users with clear selection feedback.

[0053] Next, in Figure 3 Based on the illustrated embodiment, the presentation of the identification frames is further optimized to enhance practicality and user experience. Specifically, the identification frame corresponding to each member is located in the facial area of ​​the member, which helps users quickly identify and lock on to the target member. And / or, the identification frame corresponding to each member is located in a specific area above the member's head. This avoids the visual interference that may be caused by obscuring the member's face, while maintaining a close association between the identification frame and the member, facilitating user selection without obstructing the user's view.

[0054] The flexible positioning method of this embodiment enables users to easily find and select the object they want to interact with through the identification box regardless of how the position of the members in the target video changes, thereby enhancing the smoothness and interactivity of the video interaction.

[0055] exist Figure 3Building on the illustrated embodiment, this application also optimizes the dynamic performance of the identification frames to improve the matching effect between the identification frames and members, ensuring that users can easily identify and select the target member in dynamic video scenes. Specifically, the identification frame corresponding to each member moves with the position of the member on the playback page; and / or the size of the identification frame corresponding to each member is positively correlated with the size of the member on the playback page.

[0056] In this embodiment, the identification frame is no longer limited to a static position; instead, it can adjust in real time based on the member's movements and position changes in the video. Specifically, as the member moves in the target video, the identification frame moves with them, maintaining a consistent correspondence with them and ensuring that users can consistently and accurately identify and select the target member.

[0057] In some embodiments, the size of the identification box corresponding to each member is positively correlated with the size of the member in the playback page, which enables the size of the identification box to intuitively reflect the relative importance or attention of the member in the video, while also improving the accuracy and convenience of users interacting with video content. For example, when a member occupies a larger space in the picture, that is, is the main focus of the picture, the corresponding identification box will also be larger, so that users can more easily identify and interact with it; on the contrary, if the member only occupies a smaller space in the picture, the identification box will also be reduced accordingly to avoid excessive interference with the user's viewing experience. In addition, this design also helps to clearly distinguish and highlight different participants in multi-person interactive scenarios, such as live PK, so that viewers can quickly identify the objects they want to support or interact with.

[0058] To sum up, the dynamic adjustment mechanism of this embodiment ensures that the identification frame always corresponds precisely to the member. Even when the member moves frequently or the lens is stretched and zoomed, the user can accurately identify and select the target member, thereby improving the efficiency and accuracy of the interactive selection and enhancing the overall experience of video interaction.

[0059] Figure 3 The illustrated embodiment clarifies the relationship between identification frames and member positions. To ensure that identification frames accurately correspond to each member, this embodiment further explains how to generate these identification frames. Specifically, the facial information of multiple members in the target video is first identified; then, based on their facial information, the members are labeled and their identification frames are generated.

[0060] Optionally, when a user issues an interactive command, the client first uses local AI (artificial intelligence) facial recognition technology to comprehensively scan the scene in the target video and identify all facial information. This process not only quickly locates the face of each member, but also uses deep learning algorithms to analyze each member's facial features and generate a facial identifier, or identification frame, for each member.

[0061] This facial information-based identification frame generation method in this embodiment not only improves the accuracy of member identification but also provides users with a more intuitive and convenient way to interact. By clicking or selecting these identification frames, users can identify the target members with whom they want to interact, significantly improving the efficiency of video interaction and user experience.

[0062] To further enrich the generation of identification frames, this embodiment also introduces a second user terminal to obtain identification frames for members. Specifically, identification information of multiple members in the target video is obtained from the second user terminal; multiple members are labeled based on their respective identification information, and identification frames for each of the multiple members are generated.

[0063] Member identification information refers to the characteristic data or attribute set used to uniquely identify and distinguish different members in the target video. Optionally, member identification information includes key information related to the member, such as the member's facial features, name, identity identifier, role name, or other descriptive information related to the member's identity. This is intended to facilitate accurate labeling of members in the target video content.

[0064] For example, the second user terminal uses Face ID (facial recognition) technology to annotate the names of people appearing in the video to determine the identification information of each member. This identification information is then transmitted to the first user terminal, which displays complete information including the members' faces and corresponding names on the first user terminal's interface. The first user terminal then uses this identification information to accurately identify and distinguish each member in the target video, and based on this, generates a separate identification box for each member.

[0065] The solution in this embodiment fully leverages the data resources of multiple users, enriching and diversifying the sources of identification information and further enhancing the flexibility of identification frame generation. Furthermore, by introducing identification information from a second user, real-time updates and dynamic adjustments to video content can be achieved. For example, in a live broadcast scenario, viewers can see the identification information of newly joined members in real time through the first user, allowing them to better participate in the interaction.

[0066] As can be seen from the preceding embodiments, obtaining identification information of members in the target video from the second user terminal is a key step in generating member identification frames on the first user terminal. The following further describes how this identification information is generated. Specifically, the identification information for each of the multiple members on the second user terminal is generated by the second user terminal before the target video is played, based on facial images of the members who are to be included in the target video.

[0067] For example, before a user interacts with a member of a target video, the second user terminal can enter the member's role that may appear in the video into the AI ​​model by uploading multiple pictures of their own face. Using these pictures, the AI ​​model can pre-generate identification information in advance.

[0068] This embodiment significantly improves the efficiency and accuracy of video content annotation by having the second user terminal generate identification information based on the facial images of the pre-screened actors before the target video is played. Furthermore, this preprocessing approach not only provides an accurate data foundation for subsequent identification frame generation but also significantly improves recognition accuracy and process flow. Furthermore, this method of pre-generating identification information optimizes resource allocation, reduces the burden of real-time processing, and makes the overall video content presentation more stable and efficient.

[0069] Figure 3 The illustrated embodiment achieves the association of the identification frame with the member's position. Based on this, this embodiment further expands the user's interaction with video content, providing users with a richer interactive experience. Specifically, a target special effect is determined; and the target special effect is played based on the target member.

[0070] Targeted special effects refer to visual or audio effects triggered by user operation instructions during the playback of the target video, which are used to enhance the interactive experience between the user and the members in the target video.

[0071] Specifically, when a user triggers an interactive command, the user terminal determines the target member and obtains the target special effects determined by the user. Optionally, these special effects include visual or audio effects. Then, the corresponding target special effects are played at the target member's location. For example, the animation effect of the target special effect is displayed near the target member's identification box. This implementation method not only enhances the interactivity between the user and the members in the target video, but also makes the user's interactive behavior more intuitive and vivid through the visual presentation of the target special effects, thereby improving the overall viewing experience.

[0072] In other implementations, the first user terminal determines a target special effect, the second user terminal obtains the target special effect determined by the first user terminal, and plays the target special effect based on the target member.

[0073] For example, the second user terminal receives the target special effects sent by the user from the first user terminal, including the username of the member who is desired to be highlighted. Then, the second user terminal uses Face ID technology to accurately identify and locate the face position of the target member. If the face of the target member is not in the picture during the processing, this task is cached, and after the face of the target member is subsequently identified, the target special effects are displayed. In this way, the second user terminal can achieve accurate target special effects playback, thereby enhancing the attractiveness of the target video and user engagement.

[0074] Furthermore, in some embodiments, the target special effect includes a gift-giving special effect, and the playback page of the target video presents a gift panel including a variety of gifts, so that the user can determine the gift-giving special effect based on the gift panel.

[0075] like Figure 4 and Figure 5 As shown in the left picture, a gift panel is integrated into the page, which provides a variety of gift options for users to choose from, so that users can send virtual gifts to members in the video.

[0076] In some scenarios, when users want to interact with the target members in the video, such as expressing support by sending a gift, they can select a specific gift effect by operating the gift panel on the playback page. Specifically, the gift panel lists different types of gifts, each corresponding to a specific gift-giving effect. After the user selects a gift and confirms the gift, the corresponding target effect will be displayed on the playback page of the target video based on the user's selection.

[0077] This design not only provides users with an intuitive way to express their support for the members in the video, but also enhances the fun and personalized experience of interaction through rich gift-giving special effects, while also adding more interactivity and entertainment value to the video content.

[0078] In related technologies, after a user successfully selects an interaction partner and completes an interactive action (such as sending a virtual gift), the target effect often lacks clear directionality, failing to clearly indicate that the target effect was specifically sent to a specific person. This means that not only is it difficult for the user who sent the target effect to visually verify that the selected target effect was accurately delivered, but it is also difficult for the other participants in the video to perceive that the target effect was specifically sent to them. This lack of intuitive visual feedback weakens the personalization and targeting of the interaction, fails to effectively strengthen the emotional connection between the user and the other participants in the video, and fails to enhance the user's sense of engagement and satisfaction.

[0079] In order to solve the deficiency of target special effects in the related art in terms of visual feedback, this application proposes an optimized target special effects display solution. Specifically, the target special effects are played at the designated position corresponding to the target member in the play page.

[0080] In other words, the presentation of the target special effects is not random or universal, but will be precisely associated with the position of the target member on the screen. For example, if the target special effect is a gift-giving special effect, the gift animation effect will appear directly near the target member's identification box or in a specific area of ​​the screen where they are located, rather than randomly appearing elsewhere on the playback page. This precise special effect playback method allows each member to clearly know which target special effects are specifically given to them, thus avoiding embarrassment and dissatisfaction caused by misunderstandings or confusion, and improving the user's viewing experience and sense of participation.

[0081] More specifically, if there is only one target member, the target member is given an enlarged close-up display, and the target special effect is played at the designated position corresponding to the target member in the playback page; if there are multiple target members, a special effect movement trajectory is generated based on the designated positions corresponding to the multiple target members, and the target special effect is played based on the special effect movement trajectory.

[0082] In this embodiment, by dynamically adjusting the display mode of the target special effects according to the number and position of the target members, the effect of the target special effects presentation and the user experience are further optimized. Specifically, when there is only one target member, a close-up of the member will be given. Figure 6 The figure shows a schematic diagram of a playback target special effect provided by an embodiment of the present application. Figure 6 As shown, the target special effect is a "heart-shaped gift" special effect, and the target member is the member located in the middle of the play page. When the user chooses to send the "heart-shaped gift" special effect to the target member, the special effect is displayed on the chest of the target member, and the target special effect is magnified on the play page. It is understood that in some embodiments, when the target member is magnified to a certain extent, other members on the play page may be partially squeezed out of the play page. When the target member returns to normal, multiple members will be displayed simultaneously on the play page.

[0083] This close-up display method not only highlights the target member, but also allows users to more clearly see the target member's reaction when receiving the target special effect. At the same time, the target special effect will be played at the designated position corresponding to the target member, further strengthening the focus of the visual effect.

[0084] In addition, when there are multiple target members, the generation and playback of special effect movement trajectories requires comprehensive consideration of multiple factors to ensure that the target special effects can be presented in a natural and visually attractive manner. Figure 7FIG. 1 is a schematic diagram of another embodiment of the present application providing a playback target special effect. Figure 7 As shown, the starting and ending points of the target special effect are determined based on the head position of each target member in the playback page. Taking the bounce special effect as an example, the starting point of the target special effect is the target distance above the head of the target member from left to right.

[0085] In some embodiments, the target special effect is played at the designated position corresponding to the target member in the playback page, including: adding the identity information of the target member to the visual elements of the gift-giving special effect to generate a customized gift special effect containing the identity information of the target member; playing the customized gift special effect containing the identity information of the target member at the designated position corresponding to the target member in the playback page.

[0086] For example, the visual elements of the gift-giving special effect refer to the visible parts that constitute the gift-giving special effect, including animation effects, text, graphics, colors, etc. The identity information of the target member is used to identify and describe the characteristics of the target member, including nicknames, avatars, facial expressions, and other personalized information related to the target member.

[0087] For example, the target member's avatar position on the playback page is detected (e.g., coordinates X=200, Y=500) and this position is set as the designated location for the target member. The target member's identity information is then incorporated into the gift-giving special effect to generate a customized gift effect. For example, the target member's avatar may be displayed in the center of the blooming fireworks, the target member's nickname initials may be printed on the falling ribbons, and the target member's rank badge may appear at the end of the special effect. Finally, the customized gift effect is played at the aforementioned avatar position.

[0088] This demonstrates that this embodiment enables gift-giving effects to be accurately focused on target members, enhancing the ritual and targeted nature of gift-giving. Secondly, mapping the target member's identity information to the visual elements of the gift-giving effects further enhances the personalization and exclusivity of the gift effects. Finally, displaying customized gift effects based on designated locations combines precise positioning with personalized customization, not only optimizing the visual effect but also increasing user engagement and satisfaction.

[0089] In some embodiments, in order to make the motion trajectory of the target special effect more vivid, the first user terminal generates a parabolic trajectory based on the position information, and the vertex of the trajectory is set to the center position of the two rebound points, thereby forming an arc trajectory.

[0090] After generating the trajectory, the target effect's motion parameters must also be considered, such as the curvature of the parabola and the speed of the target effect along the trajectory. These parameters can be flexibly adjusted to suit different visual styles and scene requirements. For example, the curvature of the parabola determines the target effect's jump height, while the speed of movement influences the rhythm and fluidity of the target effect. By precisely controlling these parameters, the target effect can bounce sequentially between multiple target members, creating a smooth and dynamic visual effect.

[0091] Finally, based on the generated special effect movement trajectory and the set motion parameters, the corresponding target special effect is played above the head of each target member in turn. This dynamic trajectory not only covers all target members but also guides the user's gaze between multiple members through the movement of the target special effect. This maintains the display of the target special effect while taking into account the visual balance and interactivity of the multi-member scene. This special effect playback method flexibly copes with scenes with different numbers of target members, improving the display of the target special effect and enhancing the interactive experience between the user and the video content.

[0092] Combined with the above Figures 2 to 7 , describes in detail the video processing method embodiment of the present application, and the following is combined with Figure 8 , the video processing device embodiment of the present application is described in detail. It should be understood that the description of the video processing method embodiment corresponds to the description of the video processing device embodiment, so the parts not described in detail can be referred to the previous method embodiment.

[0093] Figure 8 FIG. 1 is a schematic diagram of the structure of a video processing device provided by an embodiment of the present application. Figure 8 As shown, the video processing device 80 provided in this embodiment of the application includes:

[0094] Presentation module 810 is configured to respond to a user's interactive instruction and present, on the playback page of the target video, identification boxes corresponding to multiple members in the target video, wherein the identification box corresponding to each member is associated with the position of the member in the playback page, and the interactive instruction is configured to initiate an interactive request to at least one member in the target video;

[0095] The determination module 820 is configured to determine a target member among the multiple members in response to a user's selection instruction based on the identification box.

[0096] In one embodiment of the present application, the identification frame corresponding to each member is located in the facial area of ​​the member, and / or the identification frame corresponding to each member is located in a specific area above the head of the member.

[0097] In one embodiment of the present application, the identification box corresponding to each member moves with the position of the member in the playback page; and / or, the size of the identification box corresponding to each member is positively correlated with the size of the member in the playback page.

[0098] In one embodiment of the present application, the determination module 820 is further configured to identify facial information of multiple members in the target video; mark multiple members based on their facial information, and generate identification frames for the multiple members.

[0099] In one embodiment of the present application, the determination module 820 is further used to obtain identification information of each of the multiple members in the target video from the second user terminal; mark the multiple members based on the identification information of the multiple members, and generate identification boxes for the multiple members.

[0100] In one embodiment of the present application, identification information of each of the multiple members in the second user terminal is generated by the second user terminal before the target video is played based on facial images of the members to be included in the target video.

[0101] In one embodiment of the present application, the determination module 820 is further used to determine a target special effect, where the target special effect refers to a visual effect and / or audio effect triggered during the playback of the target video; and play the target special effect based on the target member.

[0102] In one embodiment of the present application, the determination module 820 is further configured to play a target special effect at a designated position corresponding to the target member in the play page.

[0103] In one embodiment of the present application, the determination module 820 is also used to, if the number of target members is one, give the target member an enlarged close-up display, and play the target special effects at the designated position corresponding to the target member in the playback page; if the number of target members is multiple, generate a special effect movement trajectory based on the designated positions corresponding to the multiple target members, and play the target special effects based on the special effect movement trajectory.

[0104] In one embodiment of the present application, the determination module 820 is also used to add the identity information of the target member to the visual elements of the gift-giving special effect to generate a customized gift special effect containing the identity information of the target member; and play the customized gift special effect containing the identity information of the target member at the designated position corresponding to the target member in the playback page.

[0105] In one embodiment of the present application, the presentation module 810 is further used to present a gift panel including multiple gifts on the playback page of the target video so that the user can determine the gift giving special effects based on the gift panel.

[0106] In one embodiment of the present application, the presentation module 810 is also used to present selection prompt information corresponding to at least one member on the playback page of the target video when at least one member is selected based on the identification box, so that the user can determine whether to select the at least one member as the target member. The selection prompt information refers to the prompt information displayed on the playback page of the target video indicating whether the at least one member is selected.

[0107] Below, reference Figure 9 To describe the electronic device according to the embodiment of the present application. Figure 9 Shown is a schematic structural diagram of an electronic device provided by an exemplary embodiment of the present application.

[0108] like Figure 9 As shown, the electronic device 90 includes one or more processors 901 and a memory 902 .

[0109] The processor 901 may be a central processing unit (CPU) or other forms of processing units having data processing capabilities and / or instruction execution capabilities, and may control other components in the electronic device 90 to perform desired functions.

[0110] The memory 902 may include one or more computer program products, which may include various forms of computer-readable storage media, such as volatile memory and / or non-volatile memory. The volatile memory may include, for example, random access memory (RAM) and / or cache memory. The non-volatile memory may include, for example, read-only memory (ROM), a hard disk, flash memory, etc. One or more computer program instructions may be stored on the computer-readable storage medium, and the processor 901 may execute the program instructions to implement the video processing methods of the various embodiments of the present application described above and / or other desired functions. The computer-readable storage medium may also store various contents such as interactive instructions, selection instructions, and member identification boxes.

[0111] In one example, the electronic device 90 may further include an input device 903 and an output device 904 , and these components are interconnected via a bus system and / or other forms of connection mechanisms (not shown).

[0112] The input device 903 may include, for example, a keyboard, a mouse, and the like.

[0113] The output device 904 can output various information to the outside, including interactive instructions, selection instructions, member identification boxes, etc. The output device 904 can include, for example, a display, a speaker, a printer, a communication network and its connected remote output devices, etc.

[0114] Of course, to simplify, Figure 9 Only some of the components related to the present application in the electronic device 90 are shown, and components such as a bus, an input / output interface, etc. are omitted. In addition, the electronic device 90 may further include any other appropriate components according to specific application scenarios.

[0115] In addition to the above-mentioned methods and devices, an embodiment of the present application may also be a computer program product, which includes computer program instructions, which, when executed by a processor, enable the processor to execute the steps of the video processing method according to various embodiments of the present application described above in this specification.

[0116] The computer program product may be written in any combination of one or more programming languages ​​to implement the program code for performing the operations of the embodiments of the present application, including object-oriented programming languages ​​such as Java, C++, and conventional procedural programming languages ​​such as "C" or similar programming languages. The program code may be executed entirely on the user's computing device, partially on the user's computing device, as a standalone software package, partially on the user's computing device and partially on a remote computing device, or entirely on a remote computing device or server.

[0117] In addition, an embodiment of the present application may also be a computer-readable storage medium having computer program instructions stored thereon, which, when executed by a processor, enables the processor to execute the steps of the video processing method according to various embodiments of the present application described above in this specification.

[0118] The computer-readable storage medium may be any combination of one or more readable media. The readable medium may be a readable signal medium or a readable storage medium. The readable storage medium may include, for example, but is not limited to, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, device, or device, or any combination thereof. More specific examples (a non-exhaustive list) of readable storage media include: an electrical connection having one or more wires, a portable disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination thereof.

[0119] The basic principles of the present application have been described above in conjunction with specific embodiments. However, it should be noted that the advantages, strengths, and effects mentioned in this application are merely illustrative and not restrictive, and it should not be assumed that these advantages, strengths, and effects are required of each embodiment of this application. In addition, the specific details disclosed above are merely illustrative and facilitating understanding, and are not restrictive. The above details do not limit this application to necessarily being implemented using the above specific details.

[0120] The block diagrams of the devices, devices, equipment, and systems involved in this application are merely illustrative examples and are not intended to require or imply that they must be connected, arranged, or configured in the manner shown in the block diagrams. As will be appreciated by those skilled in the art, these devices, devices, equipment, and systems can be connected, arranged, or configured in any manner. Words such as "include," "comprise," "have," and the like are open-ended words, meaning "including but not limited to," and can be used interchangeably therewith. The words "or" and "and" used herein refer to the words "and / or" and can be used interchangeably therewith, unless the context clearly indicates otherwise. The word "such as" used herein refers to the phrase "such as but not limited to," and can be used interchangeably therewith.

[0121] It should also be noted that in the apparatus, device, and method of the present application, each component or each step can be decomposed and / or recombined, and such decomposition and / or recombination should be regarded as equivalent solutions of the present application.

[0122] The above description of the disclosed aspects is provided to enable any person skilled in the art to make or use the present application. Various modifications to these aspects will be readily apparent to those skilled in the art, and the general principles defined herein may be applied to other aspects without departing from the scope of the present application. Therefore, the present application is not intended to be limited to the aspects shown herein, but rather to be accorded the widest scope consistent with the principles and novel features disclosed herein.

[0123] The above description has been provided for the purpose of illustration and description. Furthermore, this description is not intended to limit the embodiments of the present application to the forms disclosed herein. Although a number of example aspects and embodiments have been discussed above, those skilled in the art will recognize certain variations, modifications, alterations, additions, and sub-combinations thereof.

Claims

1. A video processing method, characterized in that: Applied to a first user terminal, the method includes: In response to the interaction instruction, on the playback page of the target video, identification boxes corresponding to multiple members in the target video are presented, wherein the identification box corresponding to each member is associated with the position of the member in the playback page, and the members appear at different positions, and the identification box corresponding to each member moves with the position of the member in the playback page. The interaction instruction is used to initiate an interaction request to at least one member in the target video; In response to a selection instruction based on an identification box, a target member among the multiple members is determined, wherein clicking the identification box corresponding to a member selects the member as the target member, and the target member includes a member interacting with the user of the first user terminal.

2. The video processing method according to claim 1, wherein: The identification frame corresponding to each member is located in the facial area of ​​the member, and / or, The identification frame corresponding to each member is located in a specific area above the head of the member.

3. The video processing method according to claim 1, wherein: The size of the identification frame corresponding to each member is positively correlated with the size of the member in the play page.

4. The video processing method according to claim 1, wherein: Also includes: Identifying facial information of each of the plurality of members in the target video; The plurality of members are labeled based on their respective facial information, and identification frames are generated for the plurality of members.

5. The video processing method according to claim 1, wherein: Also includes: Obtaining identification information of each of the plurality of members in the target video from the second user terminal; The multiple members are labeled based on their respective identification information, and identification boxes are generated for the multiple members.

6. The video processing method according to claim 5, characterized in that: The identification information of each of the multiple members in the second user terminal is generated by the second user terminal before the target video is played based on the facial images of the members to be included in the target video.

7. The video processing method according to any one of claims 1 to 6, characterized in that: Also includes: Determining a target special effect, where the target special effect refers to a visual effect and / or audio effect triggered during playback of the target video; The target special effect is played based on the target member.

8. The video processing method according to claim 7, wherein: Playing the target special effect based on the target member includes: The target special effect is played at the designated position corresponding to the target member in the play page.

9. The video processing method according to claim 8, characterized in that: Playing the target special effect at the designated position corresponding to the target member in the play page includes: If the number of the target member is one, the target member is given an enlarged close-up display, and the target special effect is played at the designated position corresponding to the target member in the play page; If there are multiple target members, a special effect movement trajectory is generated based on the designated positions corresponding to the multiple target members, and the target special effect is played based on the special effect movement trajectory.

10. The video processing method according to claim 8, characterized in that: The target special effect includes a gift giving special effect, and the target special effect is played at the designated position corresponding to the target member in the play page, including: Adding the identity information of the target member to the visual elements of the gift-giving special effect to generate a customized gift special effect containing the identity information of the target member; At the designated position corresponding to the target member in the play page, a customized gift special effect containing the identity information of the target member is played.

11. The video processing method according to claim 7, wherein: The target special effect includes a gift-giving special effect, and the method further includes: A gift panel including multiple gifts is presented on the playback page of the target video so that the user can determine the gift-giving special effect based on the gift panel.

12. The video processing method according to any one of claims 1 to 6, characterized in that: Before determining the target member among the multiple members in response to the selection instruction based on the identification box, the method further includes: In the case of selecting at least one member based on the identification box, selection prompt information corresponding to the at least one member is presented on the playback page of the target video to determine whether the at least one member is selected as the target member. The selection prompt information refers to the prompt information displayed on the playback page of the target video on whether the at least one member is selected.

13. A video processing device, characterized in that: Applied to a first user terminal, the device includes: a presentation module configured to present, on a playback page of a target video, identification boxes corresponding to respective members of the target video in response to an interaction instruction, wherein the identification box corresponding to each member is associated with the position of the member on the playback page, the members appearing at different positions, and the identification box corresponding to each member moving with the position of the member on the playback page, wherein the interaction instruction is configured to initiate an interaction request to at least one member of the target video; A determination module is used to determine a target member among the multiple members in response to a selection instruction based on an identification box, wherein clicking the identification box corresponding to a member selects the member as the target member, and the target member includes a member who interacts with the user of the first user terminal.

14. An electronic device, characterized in that: include: processor; as well as A memory, wherein computer program instructions are stored in the memory, and when the computer program instructions are executed by the processor, the processor is caused to perform the video processing method according to any one of claims 1 to 12.

15. A computer-readable storage medium, characterized in that The computer-readable storage medium stores computer program instructions, which, when executed by a processor, enable the processor to perform the video processing method according to any one of claims 1 to 12.

16. A computer program product comprising a computer program, characterized in that When the computer program is executed by a processor, the video processing method according to any one of claims 1 to 12 is implemented.

Citation Information

Patent Citations

  • Virtual gift giving method and device, equipment and storage medium

    CN111147877A

  • Video processing method, playing terminal and computer readable storage medium

    CN113766297A