Page display method and device, equipment and storage medium
Patent Information
- Application Number
- CN202510397735.0
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- Filing Date
- 2025-03-31
- Publication Date
- 2026-09-22
- Estimated Expiration
- 2045-03-31
AI Technical Summary
[0004]本公开提供一种页面展示方法、装置、设备及存储介质,以至少解决相关技术中视频页面展示内容单一的问题
[0134]本公开在视频播放页面中展示当前播放对象的链接标识以及所述链接标识对应的目标动态指引贴图;所述目标动态指引贴图基于图像指引信息、所述当前播放对象的资源调整文本中的至少一项生成;所述目标动态指引贴图用于引导当前用户账户与所述链接标识进行交互,所述当前用户账户为浏览所述当前播放对象的账户;本公开在页面中展示了链接标识对应的提示性动态指引贴图,提高了页面展示内容的多样性,且使得链接标识更加醒目;响应于所述当前用户账户对所述链接标识的触发操作,展示所述当前播放对象的对象详情页。本公开在页面中展示了链接标识对应的目标动态指引贴图,从而便于提示用户对链接标识进行操作,进入对象详情页;提高了用户与链接标识的交互效率。
Smart Images

Figure CN120264089B_ABST
Abstract
Description
Technical Field
[0001] This disclosure relates to the field of computer technology, and in particular to a page display method, apparatus, device, and storage medium. Background Technology
[0002] With the popularization of mobile internet and the continuous growth of user demand, more and more people are watching short videos through mobile phones and other devices. Consumers are increasingly inclined to learn about product features and usage methods through video content.
[0003] In related technologies, the production and release of a complete video involves steps such as manual product selection, material shooting, content production, and video publishing. The video production cycle is long, and the content displayed on the video page is of a single format. Summary of the Invention
[0004] This disclosure provides a page display method, apparatus, device, and storage medium to at least solve the problem of limited content displayed on video pages in related technologies. The technical solution of this disclosure is as follows:
[0005] According to a first aspect of the present disclosure, a page display method is provided, including:
[0006] The video playback page displays a link identifier for the currently playing object and a corresponding target dynamic guide image. The target dynamic guide image is generated based on at least one of image guidance information and resource adjustment text of the currently playing object. The target dynamic guide image is used to guide the current user account to interact with the link identifier, where the current user account is the account that browses the currently playing object.
[0007] In response to the current user account's trigger operation on the link identifier, the object details page of the currently playing object is displayed.
[0008] In one exemplary embodiment, the target dynamic guide image is an image that moves back and forth between a first position and a second position on the video playback page, or an image that is continuously reduced in size and then restored to its original state; the first position and the second position are different positions close to the link identifier.
[0009] In one exemplary embodiment, the target dynamic guide map includes at least one of the following:
[0010] A dynamic gesture pointing to the link identifier;
[0011] The resource adjustment text of the currently playing object is processed using at least two text styles to generate a dynamic text image;
[0012] A dynamic composite image is generated by adjusting the resource text of the currently playing object and using preset prompt images.
[0013] The dynamic texture of the prompt message style is generated by adjusting the text using the resources of the currently playing object.
[0014] In one exemplary embodiment, the target dynamic guide map is constructed based on a first map style in a first map style library, the first map style library being used to store dynamic maps whose interactive data satisfies a first filtering condition. The method further includes:
[0015] At each first preset time interval, obtain the first current interaction data of each first texture style in the first texture style library;
[0016] The first texture style that does not meet the first filtering condition in the first current interaction data is deleted from the first texture style library to obtain an updated first texture style library; the updated first texture style library is used to construct the updated video.
[0017] In one exemplary embodiment, the video playback page also displays a target user-generated content sticker of the currently playing object. The target user-generated content sticker is constructed based on a second sticker style from a second sticker style library. The second sticker style library stores user-generated content sticker styles whose interactive data satisfies a second filtering condition. The method further includes:
[0018] At every second preset time interval, obtain the second current interaction data for each second texture style in the second texture style library;
[0019] The second texture style that does not meet the second filtering condition in the second current interaction data is deleted from the second texture style library to obtain an updated second texture style library; the updated second texture style library is used to construct the updated video.
[0020] In one exemplary implementation, the video playback page is a page within the target video of the currently playing object, and the method for constructing the target video includes:
[0021] From the historical videos corresponding to the currently playing object, select videos whose popularity meets preset conditions as the selected videos;
[0022] Obtain the target material from the preset material library that matches the currently playing object in the filtered video;
[0023] The target video is constructed based on the target material and the details of the currently playing object.
[0024] In one exemplary implementation, constructing the target video based on the target material and the details of the currently playing object includes:
[0025] The video background information of the target video is synthesized based on the target material;
[0026] Parse the attribute information of the currently playing object in the filtered video, and determine the video element corresponding to the attribute information;
[0027] Based on the attribute information, the video element is filled with content to generate a video fill element;
[0028] The target video is constructed based on the video background information, the video fill elements, and the details of the currently playing object.
[0029] In one exemplary implementation, constructing the target video based on the video background information, the video fill element, and the details of the currently playing object includes:
[0030] The details of the currently playing object are broken down, and target explanatory text is generated;
[0031] The target voice style of the narration text is determined based on the geographical location of the currently playing object;
[0032] The target explanatory text and the target voice style are merged to generate target voice subtitles;
[0033] The target video is constructed based on the video background information, the video fill elements, and the target audio subtitles.
[0034] In one exemplary embodiment, the method for constructing the preset material library includes:
[0035] Obtain the key object image of the currently playing object to obtain the first material;
[0036] Obtain the object detail image of the currently playing object and the associated object image that matches the current season to obtain the second material;
[0037] Obtain the resource adjustment information of the currently playing object to obtain the third material;
[0038] A fourth source material is obtained by acquiring a composite video of the interaction between a preset user and the currently playing object.
[0039] In one exemplary embodiment, the method further includes:
[0040] Construct hotspot descriptive text corresponding to the key object image to obtain the explanatory text of the first material; the hotspot descriptive text represents descriptive text with a level of attention greater than a first threshold;
[0041] Obtain the emotional sound effect that matches the key object image to obtain the first sound effect corresponding to the first material;
[0042] Construct the relationship between the first material, the explanatory text of the first material, and the first sound effect.
[0043] In one exemplary embodiment, the method further includes:
[0044] Construct the interaction result information corresponding to the current playback object and the associated object respectively to obtain the explanatory text of the second material;
[0045] Obtain popular sound effects from preset applications and use them as the second sound effects corresponding to the second material; the popular sound effects are sound effects whose attention level is greater than a second threshold.
[0046] Construct the relationship between the second material, the explanatory text of the second material, and the second sound effect.
[0047] In one exemplary embodiment, the method further includes:
[0048] Obtain the resource information before and after adjustment corresponding to the resource adjustment information of the currently playing object;
[0049] Based on the resource information before and after the adjustment, the explanatory text for the third material is generated;
[0050] Acquire sound effects with a rhythm intensity greater than a preset intensity threshold, and use them as the third sound effects corresponding to the third material;
[0051] Construct the relationship between the third material, the explanatory text of the third material, and the third sound effect.
[0052] In one exemplary embodiment, the method further includes:
[0053] Construct interactive description text for the interaction between the preset user and the currently playing object to obtain the explanatory text of the fourth material;
[0054] Obtain the geographical location of the currently playing object and generate a voice style that matches the geographical location as the fourth sound effect corresponding to the fourth material;
[0055] Construct the relationship between the fourth material, the explanatory text of the fourth material, and the fourth sound effect.
[0056] In one exemplary implementation, constructing the target video based on the target material and the details of the currently playing object includes:
[0057] Obtain the target material category corresponding to the target material; the target material category is the first material, the second material, the third material, or the fourth material;
[0058] Based on the details of the currently playing object, extract the target prompt words for the currently playing object;
[0059] Based on the target prompts and the target associations corresponding to the target material categories, determine the target explanatory text and target sound effects corresponding to the target materials;
[0060] The target video is constructed based on the target materials, the target explanatory text, and the target sound effects.
[0061] In one exemplary embodiment, the method further includes:
[0062] Based on the relationships between the first, second, third, and fourth materials, a pre-defined corpus is constructed.
[0063] Input the target prompt words into the copywriting generation model to extract the opening video elements, middle video elements, and ending video elements;
[0064] Search the preset corpus for the explanatory text corresponding to the opening video element, the middle video element, and the ending video element;
[0065] The explanatory texts corresponding to the opening video element, the middle video element, and the closing video element are merged and processed to output the target explanatory text.
[0066] According to a second aspect of the present disclosure, a page display device is provided, comprising:
[0067] The target image display module is configured to display a link identifier of the currently playing object and a target dynamic guide image corresponding to the link identifier on the video playback page; the target dynamic guide image is generated based on at least one of image guidance information and resource adjustment text of the currently playing object; the target dynamic guide image is used to guide the current user account to interact with the link identifier, and the current user account is the account that browses the currently playing object;
[0068] The object details page display module is configured to perform a trigger operation in response to the current user account on the link identifier, and display the object details page of the currently playing object.
[0069] In one exemplary embodiment, the target dynamic guide image is an image that moves back and forth between a first position and a second position on the video playback page, or an image that is continuously reduced in size and then restored to its original state; the first position and the second position are different positions close to the link identifier.
[0070] In one exemplary embodiment, the target dynamic guide map includes at least one of the following:
[0071] A dynamic gesture pointing to the link identifier;
[0072] The resource adjustment text of the currently playing object is processed using at least two text styles to generate a dynamic text image;
[0073] A dynamic composite image is generated by adjusting the resource text of the currently playing object and using preset prompt images.
[0074] The dynamic texture of the prompt message style is generated by adjusting the text using the resources of the currently playing object.
[0075] In one exemplary embodiment, the target dynamic guide map is constructed based on a first map style in a first map style library, the first map style library being used to store dynamic maps whose interactive data satisfies a first filtering condition, and the device further includes:
[0076] The first interactive data acquisition module is configured to acquire the first current interactive data of each first texture style in the first texture style library at intervals of a first preset time.
[0077] The first update module is configured to delete the first texture style that does not meet the first filtering condition from the first texture style library to obtain an updated first texture style library; the updated first texture style library is used to construct the updated video.
[0078] In one exemplary embodiment, the video playback page further displays a target user-generated content sticker of the currently playing object. The target user-generated content sticker is constructed based on a second sticker style from a second sticker style library. The second sticker style library stores user-generated content sticker styles whose interactive data satisfies a second filtering condition. The device further includes:
[0079] The second interactive data acquisition module is configured to acquire the second current interactive data of each second texture style in the second texture style library at intervals of a second preset time.
[0080] The second update module is configured to delete the second texture style that does not meet the second filtering condition from the second texture style library to obtain an updated second texture style library; the updated second texture style library is used to construct the updated video.
[0081] In one exemplary embodiment, the video playback page is a page in the target video of the currently playing object, and the device further includes:
[0082] The video filtering module is configured to filter videos that meet preset criteria for popularity from the historical videos corresponding to the currently playing object.
[0083] The target material determination module is configured to retrieve materials from a preset material library that match the currently playing object in the filtered video, thereby obtaining the target material;
[0084] The target video construction module is configured to construct the target video based on the target material and the details of the currently playing object.
[0085] In one exemplary implementation, the target video construction module includes:
[0086] The background compositing unit is configured to perform the task of compositing video background information of the target video based on the target material;
[0087] The video element determination unit is configured to parse the attribute information of the currently playing object in the filtered video and determine the video element corresponding to the attribute information;
[0088] The fill unit is configured to perform content filling on the video element based on the attribute information to generate a video fill element;
[0089] The target video construction unit is configured to construct the target video based on the video background information, the video fill elements, and the details of the currently playing object.
[0090] In one exemplary embodiment, the target video construction unit includes:
[0091] The text generation subunit is configured to break down the details of the currently playing object and generate target explanatory text;
[0092] The target style determination subunit is configured to determine the target voice style of the narration text based on the geographical location of the currently playing object.
[0093] The target subtitle generation subunit is configured to perform a fusion of the target explanatory text and the target speech style to generate target speech subtitles;
[0094] The target video construction subunit is configured to construct the target video based on the video background information, the video fill elements, and the target audio subtitles.
[0095] In one exemplary embodiment, the apparatus further includes:
[0096] The first material acquisition module is configured to acquire the key object image of the currently playing object to obtain the first material.
[0097] The second material acquisition module is configured to acquire the object detail image of the currently playing object and the associated object image that matches the current season, thereby obtaining the second material;
[0098] The third material acquisition module is configured to acquire the resource adjustment information of the currently playing object to obtain the third material.
[0099] The fourth material acquisition module is configured to acquire a comprehensive video of the interaction between a preset user and the currently playing object, thereby obtaining the fourth material.
[0100] In one exemplary embodiment, the apparatus further includes:
[0101] The first copywriting construction module is configured to construct the hot topic description text corresponding to the key object image to obtain the explanatory copy of the first material; the hot topic description text represents the description text whose attention level is greater than a first threshold.
[0102] The first sound effect acquisition module is configured to acquire the emotional sound effect that matches the key object image, thereby obtaining the first sound effect corresponding to the first material;
[0103] The first relationship building module is configured to build the relationship between the first material, the explanatory text of the first material, and the first sound effect.
[0104] In one exemplary embodiment, the apparatus further includes:
[0105] The second copywriting construction module is configured to construct the interaction result information corresponding to the current playback object and the associated object respectively, so as to obtain the explanatory copywriting of the second material;
[0106] The second sound effect acquisition module is configured to acquire popular sound effects in a preset application as the second sound effect corresponding to the second material; the popular sound effect is a sound effect whose attention level is greater than a second threshold.
[0107] The second relationship building module is configured to build the relationship between the second material, the explanatory text of the second material, and the second sound effect.
[0108] In one exemplary embodiment, the apparatus further includes:
[0109] The adjustment information acquisition module is configured to acquire the resource information before and after adjustment corresponding to the resource adjustment information of the current playback object.
[0110] The third copywriting construction module is configured to generate explanatory copy for the third material based on the resource information before adjustment and the resource information after adjustment.
[0111] The third sound effect acquisition module is configured to acquire sound effects whose rhythm intensity is greater than a preset intensity threshold, and use them as the third sound effect corresponding to the third material.
[0112] The third relationship building module is configured to build the relationship between the third material, the explanatory text of the third material, and the third sound effect.
[0113] In one exemplary embodiment, the apparatus further includes:
[0114] The fourth copywriting construction module is configured to construct interactive description text for the interaction between the preset user and the currently playing object, thereby obtaining the explanatory text for the fourth material;
[0115] The fourth sound effect acquisition module is configured to acquire the geographical location of the currently playing object and generate a voice style that matches the geographical location as the fourth sound effect corresponding to the fourth material;
[0116] The fourth relationship building module is configured to build the relationship between the fourth material, the explanatory text of the fourth material, and the fourth sound effect.
[0117] In one exemplary implementation, the target video construction module includes:
[0118] The target material category acquisition unit is configured to acquire the target material category corresponding to the target material; the target material category is the first material, the second material, the third material, or the fourth material;
[0119] The target prompt word extraction unit is configured to extract the target prompt words of the currently playing object based on the details of the currently playing object;
[0120] The target information determination unit is configured to determine the target explanatory text and target sound effects corresponding to the target material based on the target prompt words and the target association relationship corresponding to the target material category;
[0121] The target video construction unit is configured to construct the target video based on the target material, the target narration text, and the target sound effects.
[0122] In one exemplary embodiment, the apparatus further includes:
[0123] The preset corpus construction module is configured to construct a preset corpus based on the respective relationships between the first material, the second material, the third material, and the fourth material.
[0124] The video element extraction module is configured to execute the target prompt words input into the copywriting generation model to extract the beginning video elements, the middle video elements, and the end video elements.
[0125] The script acquisition module is configured to search for the scripts corresponding to the opening video element, the middle video element, and the ending video element in the preset corpus.
[0126] The target text generation module is configured to perform fusion processing on the explanatory texts corresponding to the opening video element, the middle video element, and the closing video element, and output the target explanatory text.
[0127] According to a third aspect of the present disclosure, an electronic device is provided, comprising:
[0128] processor;
[0129] Memory used to store the processor's executable instructions;
[0130] The processor is configured to execute the instructions to implement the page display method described above.
[0131] According to a fourth aspect of the present disclosure, a computer-readable storage medium is provided that, when instructions in the computer-readable storage medium are executed by an electronic device processor, enables the electronic device to perform the page display method as described above.
[0132] According to a fifth aspect of the present disclosure, a computer program product is provided, including a computer program that, when executed by a processor, implements the page display method described above.
[0133] The technical solutions provided by the embodiments of this disclosure have at least the following beneficial effects:
[0134] This disclosure displays a link identifier for the currently playing object and a corresponding dynamic guidance image on the video playback page. The dynamic guidance image is generated based on at least one of image guidance information and resource adjustment text for the currently playing object. The dynamic guidance image guides the current user account to interact with the link identifier, where the current user account is the account browsing the currently playing object. This disclosure displays a prompting dynamic guidance image corresponding to the link identifier on the page, increasing the diversity of the displayed content and making the link identifier more prominent. In response to the current user account's trigger operation on the link identifier, the object details page of the currently playing object is displayed. This disclosure displays the target dynamic guidance image corresponding to the link identifier on the page, thereby facilitating prompts for users to interact with the link identifier and enter the object details page; improving the efficiency of user interaction with the link identifier.
[0135] It should be understood that the above general description and the following detailed description are exemplary and explanatory only, and are not intended to limit this disclosure. Attached Figure Description
[0136] The accompanying drawings, which are incorporated in and form part of this specification, illustrate embodiments consistent with this disclosure and, together with the description, serve to explain the principles of this disclosure, and are not intended to unduly limit this disclosure.
[0137] Figure 1 This is an application environment diagram illustrating a page display method according to an exemplary embodiment.
[0138] Figure 2 This is a flowchart illustrating a page display method according to an exemplary embodiment.
[0139] Figure 3 This is a schematic diagram of a video playback page according to an exemplary embodiment. Figure 1 .
[0140] Figure 4 This is a schematic diagram of a video playback page according to an exemplary embodiment. Figure 2 .
[0141] Figure 5 This is a schematic diagram of a video playback page according to an exemplary embodiment. Figure 3 .
[0142] Figure 6 This is a schematic diagram of a video playback page according to an exemplary embodiment. Figure 4 .
[0143] Figure 7 This is a schematic diagram of a page illustrating a user-generated content sticker according to an exemplary embodiment.
[0144] Figure 8 This is a flowchart illustrating a method for generating target explanatory text according to an exemplary embodiment.
[0145] Figure 9 This is a block diagram illustrating a page display device according to an exemplary embodiment.
[0146] Figure 10 This is a block diagram illustrating a server according to an exemplary embodiment.
[0147] Figure 11 This is a block diagram illustrating an electronic device for displaying a webpage according to an exemplary embodiment;
[0148] The reference numerals in the figure are as follows:
[0149] Link identifier 11, target dynamic guide image 12, target user generated content image 13. Detailed Implementation
[0150] To enable those skilled in the art to better understand the technical solutions of this disclosure, the technical solutions in the embodiments of this disclosure will be clearly and completely described below with reference to the accompanying drawings.
[0151] It should be noted that the terms "first," "second," etc., used in the specification, claims, and accompanying drawings of this disclosure are used to distinguish similar objects and are not necessarily used to describe a specific order or sequence. It should be understood that such data can be interchanged where appropriate so that the embodiments of this disclosure described herein can be implemented in orders other than those illustrated or described herein. The embodiments described in the following exemplary embodiments do not represent all embodiments consistent with this disclosure. Rather, they are merely examples of apparatuses and methods consistent with some aspects of this disclosure as detailed in the appended claims.
[0152] It should be noted that the user information (including but not limited to user device information, user personal information, etc.) and data (including but not limited to data used for display, data used for analysis, etc.) involved in this disclosure are all information and data authorized by the user or fully authorized by all parties.
[0153] To enhance the diversity of content displayed on video pages, this disclosure provides a page display method, apparatus, device, and storage medium.
[0154] Please see Figure 1 The diagram illustrates an application environment for a page display method according to an exemplary embodiment. The application environment may include a server 01 and a client 02.
[0155] Specifically, in the embodiments of this specification, server 01 may include a standalone server, a distributed server, or a server cluster composed of multiple servers. It may also be a cloud server providing basic cloud computing services such as cloud services, cloud databases, cloud computing, cloud functions, cloud storage, network services, cloud communication, middleware services, domain name services, security services, CDN (Content Delivery Network), and big data and artificial intelligence platforms. Server 01 may include a network communication unit, a processor, and a memory, etc. Specifically, server 01 can be used to generate the target video of the currently playing object and send the target video to client 02.
[0156] Specifically, in this embodiment of the specification, the client 02 may include physical devices such as smartphones, desktop computers, tablets, laptops, digital assistants, smart wearable devices, and in-vehicle terminals, and may also include software running on the physical device, such as web pages provided to users by some service providers, or applications provided to users by these service providers. Specifically, the client 02 can be used to display a playback page for the target video.
[0157] Figure 2 This is a flowchart illustrating a page display method according to an exemplary embodiment, such as... Figure 2 As shown, this method can be applied to Figure 1 The client 02 shown includes the following steps.
[0158] In step S201, the link identifier of the currently playing object and the target dynamic guide image corresponding to the link identifier are displayed on the video playback page; the target dynamic guide image is generated based on at least one of image guidance information and resource adjustment text of the currently playing object; the target dynamic guide image is used to guide the current user account to interact with the link identifier, and the current user account is the account that browses the currently playing object.
[0159] In this embodiment, the video playback page can be a page in applications such as video, shopping, and entertainment. The currently playing object can be, but is not limited to, objects such as catering, clothing, and tourist attractions. The link identifier of the currently playing object can be a text or image identifier, which may include the store name and location information of the store where the object is located. This link identifier can be used to jump to the object's details page. The target dynamic guide image is set close to the link identifier of the currently playing object and can be used to point to the link identifier. For example, the target dynamic guide image can be set above the link identifier. The target dynamic guide image can include at least one of target image and target text. The image guidance information can include, but is not limited to, gesture guidance information, directional indicator guidance information (such as arrows), etc. The target image can be the image corresponding to the image guidance information, such as a gesture guidance image or an arrow guidance image. The target text can include, but is not limited to, the resource adjustment text of the currently playing object. In this embodiment, the target dynamic guide image is an image that moves regularly within a preset range and at a preset frequency on the page. The target dynamic guide image will not cover the link identifier during dynamic movement.
[0160] In step S203, in response to the current user account's trigger operation on the link identifier, the object details page of the currently playing object is displayed.
[0161] In this embodiment, the current user account can trigger operations such as clicking, swiping, and dragging the link identifier to jump to the object details page of the currently playing object. The object details page can be the e-commerce store page of the currently playing object, or a page displaying the order entry point for the currently playing object. The object details page displays the object attribute information, price, etc. of the currently playing object; the current user account can place an order to purchase the currently playing object through the object details page, for example, by adding the currently playing object to the current user account's virtual resource list through the object details page.
[0162] This disclosure displays a link identifier for the currently playing object and a corresponding dynamic guidance image on the video playback page. The dynamic guidance image is generated based on at least one of image guidance information and resource adjustment text for the currently playing object. The dynamic guidance image guides the current user account to interact with the link identifier, where the current user account is the account browsing the currently playing object. This disclosure displays a prompting dynamic guidance image corresponding to the link identifier on the page, increasing the diversity of the displayed content and making the link identifier more prominent. In response to the current user account's trigger operation on the link identifier, the object details page of the currently playing object is displayed. This disclosure displays the target dynamic guidance image corresponding to the link identifier on the page, thereby facilitating prompts for users to interact with the link identifier and enter the object details page; improving the efficiency of user interaction with the link identifier.
[0163] In this embodiment of the disclosure, the target dynamic guide image is an image that moves back and forth between a first position and a second position on the video playback page, or an image that is continuously reduced in size and then restored to its original state; the first position and the second position are different positions close to the link identifier.
[0164] In this embodiment, the first position and the second position can be located on the same side of the link identifier to avoid the target dynamic guide image obscuring the link identifier during movement; for example, the first position and the second position can be located above the link identifier; and the first position can be above the second position. The size of the target dynamic guide image can be set according to the actual situation. During the movement, the size of the target dynamic guide image can remain unchanged, or it can continuously shrink and then return to its original state, and continuously change periodically; this embodiment sets a dynamic guide image on the page, which serves as a prompting information for the link identifier, not only improving the diversity of the page content displayed, but also making it easier to remind users of the location of the link identifier, thus improving the interaction efficiency between the user and the link identifier.
[0165] In this embodiment of the disclosure, the target dynamic guide map includes at least one of the following:
[0166] A dynamic gesture pointing to the link identifier;
[0167] The resource adjustment text of the currently playing object is processed using at least two text styles to generate a dynamic text image;
[0168] A dynamic composite image is generated by adjusting the resource text of the currently playing object and using preset prompt images.
[0169] The dynamic texture of the prompt message style is generated by adjusting the text using the resources of the currently playing object.
[0170] In this embodiment, the target dynamic guide image can be generated based on at least one of text and image; for example, the target dynamic guide image may include at least one of dynamic gesture image, dynamic text image, dynamic composite image, and dynamic image. This embodiment provides multiple styles of target dynamic guide images, improving the diversity of image display styles.
[0171] In this embodiment of the disclosure, such as Figure 3 As shown, Figure 3 This is an illustration of a video playback page. Figure 1The playback page displays the video image of the currently playing object, and also displays follow, like, comment, favorite, and share controls. It also displays a link identifier 11 and a target dynamic guide sticker 12. The target dynamic guide sticker 12 is a dynamic gesture image pointing to the link identifier, which moves back and forth continuously along the direction indicated by the double arrows on the page, to prompt the user to trigger the click on the link identifier 11.
[0172] In this embodiment of the disclosure, such as Figure 4 As shown, Figure 4 This is an illustration of a video playback page. Figure 2 The page displays a link identifier 11 and a target dynamic guide image 12. The target dynamic guide image 12 is a dynamic text image generated based on the resource adjustment text of the currently playing object. This resource adjustment text can be "offer" or "discount," and the two texts can be displayed using different text styles. The target dynamic guide image 12 corresponding to the adjustment text moves back and forth continuously along the direction indicated by the double arrows on the page to prompt the user to click on the link identifier 11.
[0173] like Figure 5 As shown, Figure 5 This is an illustration of a video playback page. Figure 3 The document displays a link identifier 11 and a target dynamic guide image 12. The target dynamic guide image 12 is a dynamic composite image, which is generated based on the resource adjustment text of the currently playing object and a preset prompt image. The resource adjustment text can be "grab a benefit". The target dynamic guide image 12 moves back and forth along the direction indicated by the double arrows on the page to prompt the user to click on the link identifier 11.
[0174] like Figure 6 As shown, Figure 6 This is an illustration of a video playback page. Figure 4 The document displays a link identifier 11 and a target dynamic guide image 12. The target dynamic guide image 12 is a prompt message-type dynamic image, generated based on the resource adjustment text of the currently playing object. This resource adjustment text could be something like "Discounts to be claimed." Figure 6 The image shows the process of the target dynamic guide map 12 shrinking and then returning to its original size, and the changes are periodic. The size can shrink until it disappears and then return to its original size, so as to prompt the user to trigger the click on the link icon 11.
[0175] In this embodiment of the disclosure, the target dynamic guide image is constructed based on a first image style in a first image style library, the first image style library being used to store dynamic images whose interactive data satisfies a first filtering condition, and the method further includes:
[0176] At each first preset time interval, obtain the first current interaction data of each first texture style in the first texture style library;
[0177] The first texture style that does not meet the first filtering condition in the first current interaction data is deleted from the first texture style library to obtain an updated first texture style library; the updated first texture style library is used to construct the updated video.
[0178] In this embodiment of the disclosure, the first preset duration can be set according to the actual situation and is not specifically limited here; a first fitting style library can be pre-built and multiple first texture styles can be stored in the first fitting style library; during the application process, a dynamic guide texture corresponding to each first texture style can be generated, and then the interaction data corresponding to each dynamic guide texture can be obtained as the first current interaction data of each first texture style; wherein, the first current interaction data may include, but is not limited to, the interaction data of the link identifier indicated by the dynamic guide texture of the first texture style.
[0179] A first data threshold corresponding to the first current interactive data can be set, and then a first filtering condition can be constructed based on the first data threshold. The first texture style that does not reach the first data threshold can be determined as a texture style that does not meet the first filtering condition and deleted from the first texture style library to update the first texture style library, thereby ensuring the business indicator requirements of the first texture style in the first texture style library.
[0180] In this embodiment of the disclosure, the video playback page also displays a target user-generated content sticker of the currently playing object. The target user-generated content sticker is constructed based on a second sticker style in a second sticker style library. The second sticker style library is used to store user-generated content sticker styles whose interactive data meets the second filtering conditions. The method further includes:
[0181] At every second preset time interval, obtain the second current interaction data for each second texture style in the second texture style library;
[0182] The second texture style that does not meet the second filtering condition in the second current interaction data is deleted from the second texture style library to obtain an updated second texture style library; the updated second texture style library is used to construct the updated video.
[0183] In this embodiment, the second preset duration may be the same as or different from the first preset duration. A second fitting style library can be pre-built, and multiple second image styles can be stored in the second fitting style library. During application, user-generated content images corresponding to each second image style can be generated, and then the interaction data corresponding to the link identifier in the page where each user-generated content image is located can be obtained as the second current interaction data for each second image style. The second current interaction data may include, but is not limited to, the interaction data of the link identifier of the second image style. A second data threshold corresponding to the second current interaction data can be set, and then a second filtering condition can be constructed based on the second data threshold. Second image styles that do not reach the second data threshold can be determined as image styles that do not meet the second filtering condition and deleted from the second image style library to update the second image style library, thereby ensuring the business indicator requirements of the second fitting styles in the second image style library.
[0184] In some embodiments, the second filtering criteria can be further adjusted based on the second current interaction data. For example, the second filtering criteria can be set to high text readability and a background color and stroke contrast higher than a preset contrast. Furthermore, the second image styles corresponding to skewed fonts and special fonts can be determined as image styles that do not meet the second filtering criteria. In some embodiments, a matching second image style can also be set for objects of specific categories; for example, for tourism-related objects, a second image style formed from text such as price discounts and attraction keywords can be set.
[0185] like Figure 7 As shown, Figure 7 This is a schematic diagram of a page displaying user-generated content stickers. It shows a link identifier 11, a target dynamic guide sticker 12, and a target user-generated content sticker 13 that meets a second filtering condition. The target dynamic guide sticker 12 is a dynamic gesture pointing to the link identifier, which moves back and forth along the direction indicated by a double arrow on the page, prompting the user to click on the link identifier 11. The text in the target user-generated content sticker 13 is legible, uses a non-skewed font, and has high contrast with the background.
[0186] In this embodiment of the disclosure, the video playback page is a page in the target video of the currently playing object, and the method for constructing the target video includes:
[0187] From the historical videos corresponding to the currently playing object, select videos whose popularity meets preset conditions as the selected videos;
[0188] Obtain the target material from the preset material library that matches the currently playing object in the filtered video;
[0189] The target video is constructed based on the target material and the details of the currently playing object.
[0190] In this embodiment, the target video may include seven elements: cover sticker, user-generated content sticker, speech synthesized using speech synthesis technology, explanatory text subtitles, background music, screen filters, and link identifier stickers. The target video can be generated based on these elements. In this embodiment, the link identifier sticker is a dynamic target guidance sticker. When creating the target video for the currently playing object, videos whose popularity meets preset conditions can be selected from its historical videos as high-quality selected videos. Alternatively, videos with business indicator data exceeding preset thresholds can be directly selected from preset applications as selected videos. The object in the selected video becomes the currently playing object. There can be one or more selected videos. By analyzing these videos, a general object explanation text structure can be abstracted, such as which elements should be included in the beginning, middle, and end of the video. Materials matching the currently playing object in the selected videos can be obtained from a preset material library, i.e., materials matching the video elements are determined to obtain target materials. Then, the target materials and the details of the currently playing object are merged to construct the target video. This embodiment further improves the content quality and diversity of the target video by merging the selected videos with the preset material library to recreate the target video and recommend it.
[0191] In this embodiment of the disclosure, constructing the target video based on the target material and the details of the currently playing object includes:
[0192] The video background information of the target video is synthesized based on the target material;
[0193] Parse the attribute information of the currently playing object in the filtered video, and determine the video element corresponding to the attribute information;
[0194] Based on the attribute information, the video element is filled with content to generate a video fill element;
[0195] The target video is constructed based on the video background information, the video fill elements, and the details of the currently playing object.
[0196] In this embodiment, there can be multiple target materials. The video background information of the target video can be synthesized from multiple target materials. Then, the attribute information of the currently playing object in the video is filtered. This attribute information may include, but is not limited to, the object's descriptive text, keywords, etc., and the corresponding video elements are determined. Video elements may include cover stickers, user-generated content images, synthesized speech, narration subtitles, background music, screen filters, link identifier images, etc. The video elements are then filled with content based on the attribute information to generate complete video fill elements. The detailed information of the currently playing object may include object images, text descriptions, etc. Finally, the video background information, video fill elements, and the detailed information of the currently playing object are fused to generate the target video. This embodiment can fill video elements based on attribute information, thereby quickly generating video fill elements and generating the target video based on the video fill elements, improving the generation efficiency and quality of the target video.
[0197] In this embodiment of the disclosure, the target video is constructed based on the video background information, the video fill elements, and the details of the currently playing object, including:
[0198] The details of the currently playing object are broken down, and target explanatory text is generated;
[0199] The target voice style of the narration text is determined based on the geographical location of the currently playing object;
[0200] The target explanatory text and the target voice style are merged to generate target voice subtitles;
[0201] The target video is constructed based on the video background information, the video fill elements, and the target audio subtitles.
[0202] In this embodiment of the disclosure, the detailed information of the currently playing object can be parsed to generate a corresponding target narration text. Then, based on the geographical location of the currently playing object, the target voice style of the narration text can be determined. Finally, based on the video background information, the video fill elements, and the target voice subtitles, a target video with geographical location characteristics can be generated.
[0203] In this embodiment of the disclosure, the method for constructing the preset material library includes:
[0204] Obtain the key object image of the currently playing object to obtain the first material;
[0205] Obtain the object detail image of the currently playing object and the associated object image that matches the current season to obtain the second material;
[0206] Obtain the resource adjustment information of the currently playing object to obtain the third material;
[0207] A fourth source material is obtained by acquiring a composite video of the interaction between a preset user and the currently playing object.
[0208] In this embodiment of the disclosure, in order to enhance the entertainment value of the video, the key object image can be a representative image selected from multiple images of the currently playing object. The key object image represents the key features of the currently playing object, which can arouse the user's interest. For example, the key object image can be a high-quality, high-contrast or high-gloss image, which can be used as the first material. For example, the key object image can be a close-up of a merchant's signature dishes or projects.
[0209] To enhance the atmosphere of the video, the second source material can be a detailed image of the currently playing object and related object images that match the current season. The detailed object image can be the featured object or a popular object, and related object images that match the current season can be determined based on the current time. For example, in summer, a restaurant can set the related object image to summer desserts. Filters that match the season can also be applied to the images to create the second source material.
[0210] To enhance the sense of discounts offered by the objects in the video, the third source material can be determined based on the resource adjustment information of the currently playing object. This resource adjustment information may include, but is not limited to, discounts or offers on individual objects or packages; it can also include highlighting effects or subtitles such as "Wow, the price has dropped!" The fourth source material can be a pre-set composite video of user interaction with the currently playing object.
[0211] In this embodiment of the disclosure, the method further includes:
[0212] Construct hotspot descriptive text corresponding to the key object image to obtain the explanatory text of the first material; the hotspot descriptive text represents descriptive text with a level of attention greater than a first threshold;
[0213] Obtain the emotional sound effect that matches the key object image to obtain the first sound effect corresponding to the first material;
[0214] Construct the relationship between the first material, the explanatory text of the first material, and the first sound effect.
[0215] In this embodiment, trending descriptive texts with a level of attention greater than a first threshold can be obtained from a preset application. The trending descriptive texts corresponding to key object images are then filtered out to obtain the first explanatory text corresponding to the first material. Alternatively, text highlighting the novelty of the group-buying product can be set as the first explanatory text, such as scene descriptions of newly opened stores, new product launches, new ways to eat, hidden features, or holiday activities. Then, emotional sound effects matching the key object images are obtained to obtain the first sound effect corresponding to the first material. For example, the first sound effect can be a sound effect used to evoke user emotions, and the explanatory text of the first material is the first explanatory text. Thus, a first association relationship is constructed between the first material, the first explanatory text, and the first sound effect, forming a comprehensive and engaging information.
[0216] In this embodiment of the disclosure, the method further includes:
[0217] Construct the interaction result information corresponding to the current playback object and the associated object respectively to obtain the explanatory text of the second material;
[0218] Obtain popular sound effects from preset applications and use them as the second sound effects corresponding to the second material; the popular sound effects are sound effects whose attention level is greater than a second threshold.
[0219] Construct the relationship between the second material, the explanatory text of the second material, and the second sound effect.
[0220] In this embodiment of the disclosure, interactive result information of positive feedback corresponding to the current playback object and the associated object can be constructed to obtain the explanatory text of the second material; for example, the taste of the food: "the meat is tender and spicy", the project experience: "I feel that all my fatigue has been dispelled", so as to achieve an immersive experience explanation; then, popular sound effects in preset applications are obtained as the second sound effects corresponding to the second material; for example, the second sound effect can be popular music with a sense of rhythm or cadence; the explanatory text of the second material is the second explanatory text; thereby constructing a second association relationship between the second material, the second explanatory text and the second sound effect, forming a comprehensive information with an atmosphere.
[0221] In this embodiment of the disclosure, the method further includes:
[0222] Obtain the resource information before and after adjustment corresponding to the resource adjustment information of the currently playing object;
[0223] Based on the resource information before and after the adjustment, the explanatory text for the third material is generated;
[0224] Acquire sound effects with a rhythm intensity greater than a preset intensity threshold, and use them as the third sound effects corresponding to the third material;
[0225] Construct the relationship between the third material, the explanatory text of the third material, and the third sound effect.
[0226] In this embodiment, the system can obtain the resource information before and after the adjustment of the resource adjustment information of the currently playing object, and form a comparison of the two types of information. The resource adjustment information may include, but is not limited to, information such as price and quantity. For example, the system can obtain the current price and the previous daily price, and form a comparison to generate a third explanatory text corresponding to the third material. Information such as the validity period of the discount can also be added to the third explanatory text. Then, a sound effect with a rhythm intensity greater than a preset intensity threshold is obtained as the third sound effect corresponding to the third material to create an atmosphere of urgency, and the discount information is emphasized. The explanatory text of the third material is the third explanatory text. Thus, a third association relationship is constructed between the third material, the third explanatory text, and the third sound effect, forming comprehensive information with a sense of discount.
[0227] In this embodiment of the disclosure, the method further includes:
[0228] Construct interactive description text for the interaction between the preset user and the currently playing object to obtain the explanatory text of the fourth material;
[0229] Obtain the geographical location of the currently playing object and generate a voice style that matches the geographical location as the fourth sound effect corresponding to the fourth material;
[0230] Construct the fourth association between the fourth material, the explanatory text of the fourth material, and the fourth sound effect.
[0231] In this embodiment of the disclosure, an interactive description text for the interaction between the preset user and the currently playing object can be constructed to obtain the explanatory text of the fourth material; then, the geographical location of the currently playing object is obtained, and a voice style matching the geographical location is generated as the fourth sound effect corresponding to the fourth material; the fourth sound effect matches the region corresponding to the geographical location; the explanatory text of the fourth material is the fourth explanatory text; thus, a fourth association relationship can be constructed between the fourth material, the fourth explanatory text, and the fourth sound effect.
[0232] In this embodiment of the disclosure, constructing the target video based on the target material and the details of the currently playing object includes:
[0233] Obtain the target material category corresponding to the target material; the target material category is the first material, the second material, the third material, or the fourth material;
[0234] Based on the details of the currently playing object, extract the target prompt words for the currently playing object;
[0235] Based on the target prompts and the target associations corresponding to the target material categories, determine the target explanatory text and target sound effects corresponding to the target materials;
[0236] The target video is constructed based on the target materials, the target explanatory text, and the target sound effects.
[0237] In this embodiment, the target material category corresponding to the target material can be the first material, the second material, the third material, or the fourth material. The target prompt words for the currently playing object can be extracted based on the details of the currently playing object. Then, based on the target prompt words and the target association relationships corresponding to the target material category, the target explanation text and target sound effects corresponding to the target material are determined. The target association relationships can include a first association relationship, a second association relationship, a third association relationship, and a fourth association relationship. For example, when the target material category is the first material, the target association relationship is the first association relationship; when the target material category is the second material, the target association relationship is the second association relationship; when the target material category is the third material, the target association relationship is the third association relationship; and when the target material category is the fourth material, the target association relationship is the fourth association relationship. The subtitles of the target explanation text can adopt various periodically changing subtitle styles; different subtitle styles can correspond to different fonts, colors, font sizes, etc. Finally, based on the determined target material, the target explanation text, and the target sound effects, a target video is constructed. During the video construction process, key information can also be set in the cover to enhance the attractiveness of the video cover.
[0238] In this embodiment of the disclosure, such as Figure 8 As shown, the method further includes:
[0239] S801: Construct a preset corpus based on the respective relationships between the first material, the second material, the third material, and the fourth material;
[0240] S803: Input the target prompt words into the copywriting generation model to extract the opening video elements, the middle video elements, and the ending video elements;
[0241] S805: Search the preset corpus for the explanatory texts corresponding to the opening video element, the middle video element, and the ending video element;
[0242] S807: Merge the explanatory texts corresponding to the opening video element, the middle video element, and the ending video element, and output the target explanatory text.
[0243] In this embodiment, a preset corpus can be constructed based on the first association relationship corresponding to the first material, the second association relationship corresponding to the second material, the third association relationship corresponding to the third material, and the fourth association relationship corresponding to the fourth material. Sample prompt words for sample objects are pre-acquired, and these sample objects are labeled with sample explanation text tags. The sample prompt words are input into a large model to extract the video elements at the beginning, middle, and end of the sample. Sample explanation text results are generated based on these video elements. The preset model is trained to obtain a text generation model based on the difference between the sample explanation text results and the sample explanation text tags. Specifically, a target loss data can be determined based on the difference between the sample explanation text results and the sample explanation text tags. Then, the model parameters of the preset model are adjusted based on the target loss data until the training termination condition is met, and the preset model at the end of training is determined as the text generation model. The training termination condition can be determined based on the target loss data and / or the number of training iterations. Then, the target prompt words are input into the text generation model to extract the opening video elements, middle video elements, and ending video elements; the corresponding explanatory texts for each of the opening video elements, middle video elements, and ending video elements are searched in the preset corpus; the explanatory texts corresponding to each of the opening video elements, middle video elements, and ending video elements are fused together to quickly and accurately output the target explanatory text, thereby improving the generation efficiency of the target video.
[0244] Figure 9 This is a block diagram illustrating a page display device according to an exemplary embodiment. (Refer to...) Figure 9 The device includes:
[0245] The target texture display module 910 is configured to display a link identifier of the currently playing object and a target dynamic guide texture corresponding to the link identifier on the video playback page; the target dynamic guide texture is generated based on at least one of image guidance information and resource adjustment text of the currently playing object; the target dynamic guide texture is used to guide the current user account to interact with the link identifier, and the current user account is the account that browses the currently playing object;
[0246] The object details page display module 920 is configured to perform a trigger operation in response to the current user account on the link identifier, and display the object details page of the currently playing object.
[0247] In one exemplary embodiment, the target dynamic guide image is an image that moves back and forth between a first position and a second position on the video playback page, or an image that is continuously reduced in size and then restored to its original state; the first position and the second position are different positions close to the link identifier.
[0248] In one exemplary embodiment, the target dynamic guide map includes at least one of the following:
[0249] A dynamic gesture pointing to the link identifier;
[0250] The resource adjustment text of the currently playing object is processed using at least two text styles to generate a dynamic text image;
[0251] A dynamic composite image is generated by adjusting the resource text of the currently playing object and using preset prompt images.
[0252] The dynamic texture of the prompt message style is generated by adjusting the text using the resources of the currently playing object.
[0253] In one exemplary embodiment, the target dynamic guide map is constructed based on a first map style in a first map style library, the first map style library being used to store dynamic maps whose interactive data satisfies a first filtering condition, and the device further includes:
[0254] The first interactive data acquisition module is configured to acquire the first current interactive data of each first texture style in the first texture style library at intervals of a first preset time.
[0255] The first update module is configured to delete the first texture style that does not meet the first filtering condition from the first texture style library to obtain an updated first texture style library; the updated first texture style library is used to construct the updated video.
[0256] In one exemplary embodiment, the video playback page further displays a target user-generated content sticker of the currently playing object. The target user-generated content sticker is constructed based on a second sticker style from a second sticker style library. The second sticker style library stores user-generated content sticker styles whose interactive data satisfies a second filtering condition. The device further includes:
[0257] The second interactive data acquisition module is configured to acquire the second current interactive data of each second texture style in the second texture style library at intervals of a second preset time.
[0258] The second update module is configured to delete the second texture style that does not meet the second filtering condition from the second texture style library to obtain an updated second texture style library; the updated second texture style library is used to construct the updated video.
[0259] In one exemplary embodiment, the video playback page is a page in the target video of the currently playing object, and the device further includes:
[0260] The video filtering module is configured to filter videos that meet preset criteria for popularity from the historical videos corresponding to the currently playing object.
[0261] The target material determination module is configured to retrieve materials from a preset material library that match the currently playing object in the filtered video, thereby obtaining the target material;
[0262] The target video construction module is configured to construct the target video based on the target material and the details of the currently playing object.
[0263] In one exemplary implementation, the target video construction module includes:
[0264] The background compositing unit is configured to perform the task of compositing video background information of the target video based on the target material;
[0265] The video element determination unit is configured to parse the attribute information of the currently playing object in the filtered video and determine the video element corresponding to the attribute information;
[0266] The fill unit is configured to perform content filling on the video element based on the attribute information to generate a video fill element;
[0267] The target video construction unit is configured to construct the target video based on the video background information, the video fill elements, and the details of the currently playing object.
[0268] In one exemplary embodiment, the target video construction unit includes:
[0269] The text generation subunit is configured to break down the details of the currently playing object and generate target explanatory text;
[0270] The target style determination subunit is configured to determine the target voice style of the narration text based on the geographical location of the currently playing object.
[0271] The target subtitle generation subunit is configured to perform a fusion of the target explanatory text and the target speech style to generate target speech subtitles;
[0272] The target video construction subunit is configured to construct the target video based on the video background information, the video fill elements, and the target audio subtitles.
[0273] In one exemplary embodiment, the apparatus further includes:
[0274] The first material acquisition module is configured to acquire the key object image of the currently playing object to obtain the first material.
[0275] The second material acquisition module is configured to acquire the object detail image of the currently playing object and the associated object image that matches the current season, thereby obtaining the second material;
[0276] The third material acquisition module is configured to acquire the resource adjustment information of the currently playing object to obtain the third material.
[0277] The fourth material acquisition module is configured to acquire a comprehensive video of the interaction between a preset user and the currently playing object, thereby obtaining the fourth material.
[0278] In one exemplary embodiment, the apparatus further includes:
[0279] The first copywriting construction module is configured to construct the hot topic description text corresponding to the key object image to obtain the explanatory copy of the first material; the hot topic description text represents the description text whose attention level is greater than a first threshold.
[0280] The first sound effect acquisition module is configured to acquire the emotional sound effect that matches the key object image, thereby obtaining the first sound effect corresponding to the first material;
[0281] The first relationship building module is configured to build the relationship between the first material, the explanatory text of the first material, and the first sound effect.
[0282] In one exemplary embodiment, the apparatus further includes:
[0283] The second copywriting construction module is configured to construct the interaction result information corresponding to the current playback object and the associated object respectively, so as to obtain the explanatory copywriting of the second material;
[0284] The second sound effect acquisition module is configured to acquire popular sound effects in a preset application as the second sound effect corresponding to the second material; the popular sound effect is a sound effect whose attention level is greater than a second threshold.
[0285] The second relationship building module is configured to build the relationship between the second material, the explanatory text of the second material, and the second sound effect.
[0286] In one exemplary embodiment, the apparatus further includes:
[0287] The adjustment information acquisition module is configured to acquire the resource information before and after adjustment corresponding to the resource adjustment information of the current playback object.
[0288] The third copywriting construction module is configured to generate explanatory copy for the third material based on the resource information before adjustment and the resource information after adjustment.
[0289] The third sound effect acquisition module is configured to acquire sound effects whose rhythm intensity is greater than a preset intensity threshold, and use them as the third sound effect corresponding to the third material.
[0290] The third relationship building module is configured to build the relationship between the third material, the explanatory text of the third material, and the third sound effect.
[0291] In one exemplary embodiment, the apparatus further includes:
[0292] The fourth copywriting construction module is configured to construct interactive description text for the interaction between the preset user and the currently playing object, thereby obtaining the explanatory text for the fourth material;
[0293] The fourth sound effect acquisition module is configured to acquire the geographical location of the currently playing object and generate a voice style that matches the geographical location as the fourth sound effect corresponding to the fourth material;
[0294] The fourth relationship building module is configured to build the relationship between the fourth material, the explanatory text of the fourth material, and the fourth sound effect.
[0295] In one exemplary implementation, the target video construction module includes:
[0296] The target material category acquisition unit is configured to acquire the target material category corresponding to the target material; the target material category is the first material, the second material, the third material, or the fourth material;
[0297] The target prompt word extraction unit is configured to extract the target prompt words of the currently playing object based on the details of the currently playing object;
[0298] The target information determination unit is configured to determine the target explanatory text and target sound effects corresponding to the target material based on the target prompt words and the target association relationship corresponding to the target material category;
[0299] The target video construction unit is configured to construct the target video based on the target material, the target narration text, and the target sound effects.
[0300] In one exemplary embodiment, the apparatus further includes:
[0301] The preset corpus construction module is configured to construct a preset corpus based on the respective relationships between the first material, the second material, the third material, and the fourth material.
[0302] The video element extraction module is configured to execute the target prompt words input into the copywriting generation model to extract the beginning video elements, the middle video elements, and the end video elements.
[0303] The script acquisition module is configured to search for the scripts corresponding to the opening video element, the middle video element, and the ending video element in the preset corpus.
[0304] The target text generation module is configured to perform fusion processing on the explanatory texts corresponding to the opening video element, the middle video element, and the closing video element, and output the target explanatory text.
[0305] Regarding the apparatus in the above embodiments, the specific manner in which each module performs its operation has been described in detail in the embodiments related to the method, and will not be elaborated upon here.
[0306] In one exemplary embodiment, an electronic device is also provided, including a processor; a memory for storing processor-executable instructions; wherein, when the processor is configured to execute the instructions stored in the memory, it implements the page display method provided in any of the above embodiments.
[0307] The electronic device can be a terminal, a server, or a similar computing device. Taking a server as an example... Figure 10 This is a block diagram illustrating an electronic device according to an exemplary embodiment, such as... Figure 10As shown, the server 1000 can vary significantly due to different configurations or performance. It may include one or more central processing units (CPUs) 1010 (CPUs 1010 may include, but are not limited to, microprocessors (MCUs) or programmable logic devices (FPGAs), a memory 1030 for storing data, and one or more storage media 1020 (e.g., one or more mass storage devices) for storing application programs 1023 or data 1022. The memory 1030 and storage media 1020 may be temporary or persistent storage. The program stored in the storage media 1020 may include one or more modules, each module may include a series of instruction operations on the server. Furthermore, the CPU 1010 may be configured to communicate with the storage media 1020 and execute the series of instruction operations in the storage media 1020 on the server 1000. Server 1000 may also include one or more power supplies 1060, one or more wired or wireless network interfaces 1050, one or more input / output interfaces 1040, and / or one or more operating systems 1021, such as Windows Server™, Mac OS X™, Unix™, Linux™, FreeBSD™, etc.
[0308] The input / output interface 1040 can be used to receive or send data via a network. Specific examples of the network described above may include a wireless network provided by the communication provider of server 1000. In one example, the input / output interface 1040 includes a network interface controller (NIC), which can connect to other network devices via a base station to communicate with the Internet. In another example, the input / output interface 1040 may be a radio frequency (RF) module for wireless communication with the Internet.
[0309] Those skilled in the art will understand that Figure 10 The structure shown is for illustrative purposes only and does not limit the structure of the aforementioned electronic device. For example, server 1000 may also include... Figure 10 The more or fewer components shown, or having the same Figure 10 The different configurations shown.
[0310] In one exemplary embodiment, a computer-readable storage medium including instructions is also provided, such as a memory 1030 including instructions, which can be executed by a processor 1010 of the device 1000 to perform the above-described method. Optionally, the computer-readable storage medium may be a ROM, random access memory (RAM), CD-ROM, magnetic tape, floppy disk, and optical data storage device, etc.
[0311] Figure 11 This is a block diagram illustrating an electronic device for displaying a webpage according to an exemplary embodiment. The electronic device may be a terminal, and its internal structure diagram may be as follows: Figure 11 As shown, the electronic device includes a processor, memory, network interface, display screen, and input devices connected via a system bus. The processor provides computing and control capabilities. The memory includes a non-volatile storage medium and internal memory. The non-volatile storage medium stores the operating system and computer programs. The internal memory provides an environment for the operation of the operating system and computer programs in the non-volatile storage medium. The network interface is used to communicate with external terminals via a network connection. When the computer program is executed by the processor, it implements a page display method. The display screen can be a liquid crystal display (LCD) or an e-ink display. The input devices can be a touch layer covering the display screen, buttons, a trackball, or a touchpad mounted on the device's casing, or an external keyboard, touchpad, or mouse.
[0312] Those skilled in the art will understand that Figure 11 The structure shown is merely a block diagram of a portion of the structure related to the present disclosure and does not constitute a limitation on the electronic device to which the present disclosure is applied. A specific electronic device may include more or fewer components than those shown in the figure, or combine certain components, or have different component arrangements.
[0313] In an exemplary embodiment, an electronic device is also provided, comprising:
[0314] A processor; a memory for storing processor-executable instructions; wherein the processor is configured to execute the instructions to implement the page display method described above.
[0315] In an exemplary embodiment, a computer-readable storage medium including instructions is also provided, such as a memory including instructions, which can be executed by a processor of an electronic device to complete the page display method described above. Optionally, the computer-readable storage medium may be a ROM, random access memory (RAM), CD-ROM, magnetic tape, floppy disk, and optical data storage device, etc.
[0316] In an exemplary embodiment, a computer program product is also provided, including a computer program that, when executed by a processor, implements the page display method described above.
[0317] This disclosure displays a link identifier for the currently playing object and a corresponding dynamic guidance image on the video playback page. The dynamic guidance image is generated based on at least one of image guidance information and resource adjustment text for the currently playing object. The dynamic guidance image guides the current user account to interact with the link identifier, where the current user account is the account browsing the currently playing object. This disclosure displays a prompting dynamic guidance image corresponding to the link identifier on the page, increasing the diversity of the displayed content and making the link identifier more prominent. In response to the current user account's trigger operation on the link identifier, the object details page of the currently playing object is displayed. This disclosure displays the target dynamic guidance image corresponding to the link identifier on the page, thereby facilitating prompts for users to interact with the link identifier and enter the object details page; improving the efficiency of user interaction with the link identifier.
[0318] Those skilled in the art will understand that all or part of the processes in the methods of the above embodiments can be implemented by a computer program instructing related hardware. This computer program can be stored in a non-volatile computer-readable storage medium. When executed, the computer program can include the processes of the embodiments of the above methods. Any references to memory, storage, databases, or other media used in the embodiments provided in this application can include non-volatile and / or volatile memory. Non-volatile memory can include read-only memory (ROM), programmable ROM (PROM), electrically programmable ROM (EPROM), electrically erasable programmable ROM (EEPROM), or flash memory. Volatile memory can include random access memory (RAM) or external cache memory. By way of illustration and not limitation, RAM is available in various forms, such as static RAM (SRAM), dynamic RAM (DRAM), synchronous DRAM (SDRAM), dual data rate SDRAM (DDRSDRAM), enhanced SDRAM (ESDRAM), synchronous link DRAM (SLDRAM), RAMbus direct RAM (RDRAM), direct memory bus dynamic RAM (DRDRAM), and RAMbus dynamic RAM (RDRAM), etc.
[0319] Other embodiments of this disclosure will readily occur to those skilled in the art upon consideration of the specification and practice of the invention disclosed herein. This application is intended to cover any variations, uses, or adaptations of this disclosure that follow the general principles of this disclosure and include common knowledge or customary techniques in the art not disclosed herein. The specification and examples are to be considered exemplary only, and the true scope and spirit of this disclosure are indicated by the following claims.
[0320] It should be understood that this disclosure is not limited to the precise structures described above and shown in the accompanying drawings, and various modifications and changes can be made without departing from its scope. The scope of this disclosure is limited only by the appended claims.
Claims
1. A page display method, characterized in that, include: From the historical videos corresponding to the currently playing object, select videos whose popularity meets the preset criteria as the filter videos; Obtain the target material from the preset material library that matches the currently playing object in the filtered video; The method for constructing the preset material library includes: acquiring key object images of the currently playing object to obtain first material; acquiring object detail images of the currently playing object and related object images matching the current season to obtain second material; acquiring resource adjustment information of the currently playing object to obtain third material; and acquiring a comprehensive video of a preset user interacting with the currently playing object to obtain fourth material. Based on the target material and the details of the currently playing object, a target video is constructed; The video playback page displays a link identifier for the currently playing object and a corresponding dynamic guide image. The dynamic guide image is generated based on at least one of image guidance information and resource adjustment text of the currently playing object. The dynamic guide image is used to guide the current user account to interact with the link identifier, where the current user account is the account that browses the currently playing object. The video playback page is the page in the target video of the currently playing object. In response to the current user account's trigger operation on the link identifier, the object details page of the currently playing object is displayed.
2. The method according to claim 1, characterized in that, The target dynamic guide image is an image that moves back and forth between a first position and a second position on the video playback page, or an image that is continuously reduced in size and then restored to its original state; the first position and the second position are different positions close to the link icon.
3. The method according to claim 2, characterized in that, The target dynamic guide texture includes at least one of the following: A dynamic gesture pointing to the link identifier; The resource adjustment text of the currently playing object is processed using at least two text styles to generate a dynamic text image; A dynamic composite image is generated by adjusting the resource text of the currently playing object and using preset prompt images. The dynamic texture of the prompt message style is generated by adjusting the text using the resources of the currently playing object.
4. The method according to claim 1, characterized in that, The target dynamic guide image is constructed based on a first image style in a first image style library, the first image style library being used to store dynamic images whose interactive data meets the first filtering condition. The method further includes: At each first preset time interval, obtain the first current interaction data of each first texture style in the first texture style library; The first texture style that does not meet the first filtering condition in the first current interaction data is deleted from the first texture style library to obtain an updated first texture style library; the updated first texture style library is used to construct the updated video.
5. The method according to claim 4, characterized in that, The video playback page also displays a target user-generated content sticker for the currently playing object. This target user-generated content sticker is constructed based on a second sticker style from a second sticker style library. The second sticker style library stores user-generated content sticker styles whose interactive data meets the second filtering criteria. The method further includes: At every second preset time interval, obtain the second current interaction data for each second texture style in the second texture style library; The second texture style that does not meet the second filtering condition in the second current interaction data is deleted from the second texture style library to obtain an updated second texture style library; the updated second texture style library is used to construct the updated video.
6. The method according to claim 1, characterized in that, The step of constructing a target video based on the target material and the details of the currently playing object includes: The video background information of the target video is synthesized based on the target material; Parse the attribute information of the currently playing object in the filtered video, and determine the video element corresponding to the attribute information; Based on the attribute information, the video element is filled with content to generate a video fill element; The target video is constructed based on the video background information, the video fill elements, and the details of the currently playing object.
7. The method according to claim 1, characterized in that, The method further includes: Obtain the resource information before and after adjustment corresponding to the resource adjustment information of the currently playing object; Based on the resource information before and after the adjustment, the explanatory text for the third material is generated; Acquire sound effects with a rhythm intensity greater than a preset intensity threshold, and use them as the third sound effects corresponding to the third material; Construct the relationship between the third material, the explanatory text of the third material, and the third sound effect.
8. The method according to claim 1, characterized in that, The step of constructing a target video based on the target material and the details of the currently playing object includes: Obtain the target material category corresponding to the target material; the target material category is the first material, the second material, the third material, or the fourth material; Based on the details of the currently playing object, extract the target prompt words for the currently playing object; Based on the target prompts and the target associations corresponding to the target material categories, determine the target explanatory text and target sound effects corresponding to the target materials; The target video is constructed based on the target materials, the target explanatory text, and the target sound effects.
9. A page display device, characterized in that, include: The video filtering module is configured to filter videos from the historical videos corresponding to the currently playing object that meet preset criteria for popularity. The target material determination module is configured to retrieve materials from a preset material library that match the currently playing object in the filtered video, thereby obtaining the target material; The method for constructing the preset material library includes: acquiring key object images of the currently playing object to obtain first material; acquiring object detail images of the currently playing object and related object images matching the current season to obtain second material; acquiring resource adjustment information of the currently playing object to obtain third material; and acquiring a comprehensive video of a preset user interacting with the currently playing object to obtain fourth material. The target video construction module is configured to construct a target video based on the target material and the details of the currently playing object; The target image display module is configured to display a link identifier of the currently playing object and a corresponding target dynamic guide image on the video playback page. The target dynamic guide image is generated based on at least one of image guidance information and resource adjustment text of the currently playing object. The target dynamic guide image is used to guide the current user account to interact with the link identifier, where the current user account is the account that browses the currently playing object. The video playback page is the page in the target video of the currently playing object. The object details page display module is configured to perform a trigger operation in response to the current user account on the link identifier, and display the object details page of the currently playing object.
10. An electronic device, characterized in that, include: processor; Memory used to store the processor's executable instructions; The processor is configured to execute the instructions to implement the page display method as described in any one of claims 1-8.
11. A computer-readable storage medium, characterized in that, When the instructions in the computer-readable storage medium are executed by an electronic device processor, the electronic device is able to perform the page display method as described in any one of claims 1-8.
12. A computer program product comprising computer instructions, characterized in that, When the computer instructions are executed by the processor, they implement the page display method according to any one of claims 1-8.
Citation Information
Patent Citations
Live broadcast room display method and device based on virtual resources, equipment and storage medium
CN118870044A
Marketing commodity short video generation method and device based on AI, medium and product
CN119130502A
Livestreaming processing method and apparatus, electronic device, and computer-readable storage medium
US20220360825A1