Content display method, apparatus, and storage medium

By acquiring users' gaze direction and gaze duration, user profiles are determined, solving the problem that existing display devices cannot provide personalized displays and achieving more efficient advertising results.

CN115082111BActive Publication Date: 2026-02-06BOE TECHNOLOGY GROUP CO LTD
View PDF 1 Cites 0 Cited by

Patent Information

Application Number
CN202210704177.4
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2022-06-21
Publication Date
2026-02-06
Estimated Expiration
2042-06-21

AI Technical Summary

Technical Problem

Existing display devices cannot personalize displays according to user needs, resulting in poor advertising effectiveness.

Method used

By obtaining the target user's gaze direction and gaze duration, a user profile is determined, including attention point information and feature information, and then the target content corresponding to the user profile is displayed on the display interface.

Benefits of technology

It enables users to view relevant content based on their interests and personal attributes, thus meeting user needs and improving the effectiveness of advertising.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN115082111B_ABST
    Figure CN115082111B_ABST
Patent Text Reader

Abstract

Embodiments of the present disclosure provide a content display method and device and a storage medium, and relate to the field of information processing, and can solve the problem of poor promotion effect in the related art. The method comprises: obtaining a gaze direction of a target user in a target range at a current time; when the gaze direction of the target user at the current time is a display interface of a display device and the duration for which the target user gazes at the display interface exceeds a preset duration, determining a user portrait of the target user; wherein the user portrait comprises at least one of a first label and a second label; the first label is used to represent the attention point information of the target user, and the second label is used to represent the feature information of the target user; and according to the user portrait, displaying target content corresponding to the user portrait on the display interface. The present disclosure can meet the actual needs of users and improve the promotion effect.
Need to check novelty before this filing date? Find Prior Art

Description

TECHNICAL FIELD

[0001] The present disclosure relates to the field of information processing, and in particular, to a content display method and device and a storage medium. BACKGROUND

[0002] In public places such as scenic spots, shopping malls, and exhibitions, display devices capable of displaying content are usually provided to display corresponding promotional content to surrounding users. However, the current display devices usually fixedly display the pre-set promotional content and cannot display personalized content according to the needs of surrounding users, resulting in poor promotional effects. SUMMARY

[0003] Embodiments of the present disclosure aim to provide a content display method, device and storage medium to meet the needs of users and improve the promotional effect.

[0004] To achieve the above-mentioned purpose, embodiments of the present disclosure provide the following technical solutions:

[0005] In one aspect, a content display method is provided, which includes: obtaining a gaze direction of a target user in a target range at a current time; determining a user portrait of the target user when the gaze direction of the target user at the current time is a display interface of a display device and the target user gazes at the display interface for a duration exceeding a preset duration; wherein the user portrait includes at least one of a first label and a second label; the first label is used to represent the attention point information of the target user, and the second label is used to represent the characteristic information of the target user; and displaying target content corresponding to the user portrait on the display interface according to the user portrait.

[0006] Based on the above technical solutions, the content display device in the present disclosure can obtain the gaze direction of the target user in the target range at the current time, and further determine the user portrait of the target user when the target user gazes at the display interface of the display device for a duration exceeding a preset duration, so as to display target content corresponding to the user portrait on the display interface according to the user portrait. Since the user portrait includes at least one of the first label used to represent the attention point information of the target user and the second label used to represent the characteristic information of the target user, the content display device in the present disclosure can display corresponding target content for the target user based on two dimensions of the attention content and the personal attributes of the target user, meet the actual needs of users, and improve the promotional effect.

[0007] In some embodiments, the method includes: determining the target content according to the user portrait; and controlling the display device to display the target content according to the content currently displayed on the display interface.

[0008] In some embodiments, the method comprises: determining, for each label in the user portrait of the target user, a weight of each content in the preset content set corresponding to the label; calculating, for each content in the preset content set, a sum of the weights of each label corresponding to the content; and taking the content in the preset content set with the largest sum of weights as the target content.

[0009] In some embodiments, the method comprises: inputting each label in the user portrait of the target user into a preset decision model to obtain a plurality of predicted contents; the preset decision model comprises a plurality of decision sub-models; each decision sub-model is configured to determine a predicted content according to each label; and taking the predicted content that is the same and has the largest quantity in the plurality of predicted contents as the target content.

[0010] In some embodiments, the method comprises: when the content displayed by the display interface is a first content, controlling the display device to display the target content after the display of the first content is completed; the first content is related to the user portrait of the user; and when the content displayed by the display interface is a second content, controlling the display device to stop displaying the second content and display the target content; the second content is not related to the user portrait of the user.

[0011] In some embodiments, the method comprises: obtaining a route trajectory of a target user; wherein the route trajectory comprises one or more regions passed by the target user in chronological order and time information of the target user reaching the one or more regions; determining a target residence region of the target user according to the route trajectory; the target residence region is a region in which the target user stays for more than a preset time length among the one or more regions; and determining a first label according to region information of the target residence region.

[0012] In some embodiments, the method comprises: obtaining face image information of a target user; determining identification information of the target user according to the face image information of the target user and a first preset database; the first preset database comprises a plurality of face image information and identification information corresponding to each face image information; and obtaining a route trajectory of the target user from a second preset database according to the identification information of the target user; the second preset database comprises identification information of a plurality of users and a route trajectory of each user.

[0013] In some embodiments, the method comprises: determining a target time length and a target distance of the target user in the route trajectory; wherein the target time length is a time difference between a time when the target user reaches a first region and a time when the target user reaches a second region; the target distance is a distance between the first region and the second region; the first region is any region passed by the target user, and the second region is a next region passed by the target user after the first region; and determining the target residence region of the target user as the first region when the target time length and the target distance satisfy a first preset condition.

[0014] In some embodiments, the method comprises: obtaining image information of the target user; the image information of the target user comprises at least one of face image information and body image information; inputting the image information of the target user into the feature information recognition model to obtain a second label of the target user.

[0015] In some embodiments, there are multiple target users in the target range, the first target label of the target user is a label that is the same and has the largest quantity in the first labels of the multiple target users, and the second target label of the target user is a label that is the same and has the largest quantity in the second labels of the multiple target users; the user portrait of the target user comprises at least one of the first target label and the second target label.

[0016] In some embodiments, the method comprises: determining, according to the second label in the user portrait, special effect information to be displayed by the display device; and the special effect information is used to render the target content.

[0017] In some embodiments, the method comprises: obtaining body image information of the target user; performing image recognition on the body image information of the target user to obtain user gesture information of the target user; and when the user gesture information of the target user is preset gesture information, controlling the display device to stop displaying the target content and display third content.

[0018] In another aspect, a content display device is provided, comprising: an acquisition unit configured to acquire a gaze direction of a target user in a target range at a current time; a processing unit configured to determine a user portrait of the target user when the gaze direction of the target user at the current time is a display interface of the display device and a duration for which the target user gazes at the display interface exceeds a preset duration; wherein the user portrait comprises at least one of a first label and a second label; the first label is used to represent attention point information of the target user, and the second label is used to represent feature information of the target user; and the processing unit is further configured to display, according to the user portrait, target content corresponding to the user portrait on the display interface.

[0019] In some embodiments, the processing unit is configured to determine, according to the user portrait, target content; and control the display device to display the target content according to content currently displayed on the display interface.

[0020] In some embodiments, the processing unit is configured to, for each label in the user portrait of the target user, determine a weight of each content corresponding to the label in a preset content set; for each content in the preset content set, calculate a sum of weights of each label corresponding to the content; and determine, as the target content, a content with the largest sum of weights in the preset content set.

[0021] In some embodiments, the processing unit is configured to input each label in the user portrait of the target user into a preset decision model to obtain a plurality of predicted contents; the preset decision model comprises a plurality of decision sub-models; each decision sub-model is used to determine a predicted content according to each label; and the same and most numerous predicted contents in the plurality of predicted contents are taken as the target content.

[0022] In some embodiments, the processing unit is configured to, when the content displayed by the display interface is first content, control the display device to display the target content after the display of the first content is completed; the first content is related to the user portrait of the user; and when the content displayed by the display interface is second content, control the display device to stop displaying the second content and display the target content; the second content is not related to the user portrait of the user.

[0023] In some embodiments, the obtaining unit is configured to obtain a route trajectory of the target user; the route trajectory comprises one or more regions passed by the target user in a time sequence and time information of the target user reaching the one or more regions; and the processing unit is configured to determine a target residence region of the target user according to the route trajectory; the target residence region is a region in which the target user stays for more than a preset time length among the one or more regions; and the processing unit is further configured to determine the first label according to region information of the target residence region.

[0024] In some embodiments, the obtaining unit is configured to obtain face image information of the target user; the processing unit is configured to determine identification information of the target user according to the face image information of the target user and a first preset database; the first preset database comprises a plurality of face image information and identification information corresponding to each face image information; and the processing unit is further configured to obtain the route trajectory of the target user from a second preset database according to the identification information of the target user; the second preset database comprises identification information of a plurality of users and a route trajectory of each user.

[0025] In some embodiments, the processing unit is configured to determine a target time length and a target distance of the target user in the route trajectory; the target time length is a time difference between a time at which the target user reaches a first region and a time at which the target user reaches a second region; the target distance is a distance between the first region and the second region; the first region is any region passed by the target user, and the second region is a next region passed by the target user after the first region; and the target residence region of the target user is determined to be the first region when the target time length and the target distance satisfy a first preset condition.

[0026] In some embodiments, the obtaining unit is configured to obtain image information of the target user; the image information of the target user comprises at least one of face image information and body image information; and the processing unit is configured to input the image information of the target user into the feature information recognition model to obtain a second label of the target user.

[0027] In some embodiments, there are multiple target users in the target range, the first target label of the target user is a label that is the same and has the largest quantity in the first labels of the multiple target users, the second target label of the target user is a label that is the same and has the largest quantity in the second labels of the multiple target users, and the user portrait of the target user comprises at least one of the first target label and the second target label.

[0028] In some embodiments, the processing unit is configured to determine, according to the second label in the user portrait, special effect information to be displayed by the display device; and the special effect information is used to render the target content.

[0029] In some embodiments, the obtaining unit is configured to obtain body image information of the target user; the processing unit is configured to perform image recognition on the body image information of the target user to obtain user gesture information of the target user; and the processing unit is further configured to control the display device to stop displaying the target content and display third content when the user gesture information of the target user is preset gesture information.

[0030] In another aspect, a computer readable storage medium is provided. The computer readable storage medium stores computer program instructions. When the computer program instructions are executed on a computer (e.g., a content display device), the computer program instructions cause the computer to perform the content display method according to any one of the above embodiments.

[0031] In yet another aspect, a computer program product is provided. The computer program product includes computer program instructions. When the computer program instructions are executed on a computer (e.g., a content display device), the computer program instructions cause the computer to perform the content display method according to any one of the above embodiments.

[0032] In yet another aspect, a computer program is provided. When the computer program is executed on a computer (e.g., a content display device), the computer program causes the computer to perform the content display method according to any one of the above embodiments.

[0033] In yet another aspect, a content display device is provided. The device includes a processor and a memory. The memory is configured to store computer programs or instructions, and the processor is configured to execute the computer programs or instructions to implement the content display method according to any one of the above embodiments.

[0034] In yet another aspect, a chip is provided, which includes a processor and a communication interface, the communication interface and the processor are coupled, and the processor is configured to run computer programs or instructions to implement the content display method according to any one of the above embodiments.

[0035] For example, the chip provided in the present disclosure further includes a memory for storing the computer programs or instructions.

[0036] It should be noted that the above computer instructions can be stored on the computer readable storage medium in whole or in part. The computer readable storage medium can be packaged together with the processor of the device or packaged separately from the processor of the device, and the present disclosure does not limit the same.

[0037] In yet another aspect, a content display system is provided, which includes a content display device, a display device and an image acquisition device, the display device is configured to display content, the image acquisition device is configured to acquire images, and the content display device is configured to execute the content display method according to any one of the above embodiments.

[0038] In the present disclosure, the name of the content display device does not limit the device or functional module itself, and in actual implementation, these devices or functional modules can appear in other names. As long as the functions of the devices or functional modules are similar to those of the present disclosure, they belong to the scope of the claims of the present disclosure and equivalent technologies. BRIEF DESCRIPTION OF DRAWINGS

[0039] In order to more clearly illustrate the technical solutions in the present disclosure, the drawings needed in some embodiments of the present disclosure will be briefly introduced below. Obviously, the drawings described below are only some of the drawings of the embodiments of the present disclosure, and other drawings can also be obtained by those skilled in the art according to these drawings. In addition, the drawings described below can be regarded as schematic diagrams, and are not limited to the actual size, actual process, actual time sequence, etc. of the products involved in the embodiments of the present disclosure.

[0040] Figure 1 An architecture diagram of a content display system according to some embodiments is provided;

[0041] Figure 2 A region diagram of a preset scene according to some embodiments is provided;

[0042] Figure 3 A scene diagram of a display device according to some embodiments is provided;

[0043] Figure 4 A structure diagram of a display device according to some embodiments is provided;

[0044] Figure 5A flowchart of a content display method according to some embodiments;

[0045] Figure 6 A scene diagram of a content display according to some embodiments;

[0046] Figure 7A A side view of a content display scene according to some embodiments;

[0047] Figure 7B A top view of a content display scene according to some embodiments;

[0048] Figure 8 A flowchart of another content display method according to some embodiments;

[0049] Figure 9 A flowchart of another content display method according to some embodiments;

[0050] Figure 10 A flowchart of another content display method according to some embodiments;

[0051] Figure 11 A structure diagram of a decision sub-model according to some embodiments;

[0052] Figure 12 A flowchart of another content display method according to some embodiments;

[0053] Figure 13 A flowchart of a trajectory route generation method according to some embodiments;

[0054] Figure 14 A flowchart of another content display method according to some embodiments;

[0055] Figure 15 A flowchart of another content display method according to some embodiments;

[0056] Figure 16 A flowchart of another content display method according to some embodiments;

[0057] Figure 17 A structure diagram of a content display device according to some embodiments;

[0058] Figure 18 A structure diagram of another content display device according to some embodiments. DETAILED DESCRIPTION

[0059] In the following, the technical solutions in the embodiments of the present disclosure will be described clearly and completely with reference to the drawings. Obviously, the described embodiments are only a part of the embodiments of the present disclosure, but not all the embodiments. Based on the embodiments provided by the present disclosure, all other embodiments obtained by a person of ordinary skill in the art belong to the scope of protection of the present disclosure.

[0060] Unless otherwise required by context, the term "comprise" and its other forms such as "comprises" and "comprising" are to be construed as open-ended, i.e. as "including, but not limited to", in the description and the claims. In the description, the terms "one embodiment", "some embodiments", "exemplary embodiments", "example", "specific example" or "some examples" are not necessarily referring to the same embodiment or example. Furthermore, the described features, structures, materials or characteristics can be combined in any suitable manner in one or more embodiments or examples.

[0061] Hereinafter, the terms "first", "second", etc. are used only for the purpose of description and should not be understood as indicating or implying relative importance or implying the number of the technical features indicated. Therefore, the features defined with "first", "second" can explicitly or implicitly include one or more of the features. In the description of the embodiments of the present disclosure, the meaning of "a plurality of" is two or more, unless otherwise specified.

[0062] In describing some embodiments, "coupled" and "connected", and their derivatives, can be used. For example, the term "connected" can be used in describing some embodiments to indicate that two or more components are in direct physical or electrical contact with each other. For another example, the term "coupled" can be used in describing some embodiments to indicate that two or more components are in direct physical or electrical contact with each other. However, the term "coupled" or "communicatively coupled" can also mean that two or more components are not in direct contact with each other, but still cooperate or interact with each other. The embodiments disclosed herein are not necessarily limited to the content herein.

[0063] "At least one of A, B, and C" has the same meaning as "at least one of A, B, or C," and includes the following combinations for A, B, and C: A only, B only, C only, a combination of A and B, a combination of A and C, a combination of B and C, and a combination of A and B and C.

[0064] "A and / or B" includes the following three combinations: A only, B only, and a combination of A and B.

[0065] As used herein, the term "if' is, optionally, interpreted as meaning "when" or "while" or "in response to a determination" or "in response to a detection of. Similarly, the phrase "if determined," or "if detected [a stated condition or event]" is, optionally, interpreted as meaning "upon a determination of" or "in response to a determination of" or "upon a detection of [a stated condition or event]" or "in response to a detection of [a stated condition or event]."

[0066] The use of "adapted to" or "configured to" herein means open and inclusive language that does not exclude devices that are adapted to or configured to perform additional tasks or steps.

[0067] Additionally, the use of "based on" means open and inclusive, as a process, step, calculation, or other action "based on" one or more stated conditions or values can in practice be based on additional conditions or values beyond those stated.

[0068] As used herein, "about," "approximately," or "circa" includes the recited value and the average value within an acceptable range of deviation from the particular value, as determined by one of ordinary skill in the art considering the measurement in question and the error in measuring the particular quantity (i.e., the limitations of the measurement system).

[0069] In public places such as scenic spots, shopping malls, exhibitions, and the like, display devices capable of displaying content are usually provided to display corresponding promotional content to surrounding users.

[0070] Currently, the display device in the related art displays promotional content in a fixed order by pre-setting the promotional content to be displayed, to achieve the effect of promoting corresponding content to surrounding users. However, this scheme cannot recommend content according to the interests of users, and does not meet the needs of users, so the promotional effect is poor.

[0071] In view of this, this disclosure provides a content display method that determines a user profile of a user gazing at a display device, and displays target content corresponding to the user profile on the display interface of the display device according to the user profile. Since the determined user profile includes at least one of a first tag representing the user's focus information and a second tag representing the user's characteristic information, this disclosure can display corresponding content for the user based on two dimensions: the user's focus content and personal attributes, thus meeting the user's needs and improving the promotional effect.

[0072] The embodiments of this disclosure will now be described in detail with reference to the accompanying drawings.

[0073] Figure 1 This is an architecture diagram of a content display system 10 provided according to some embodiments, such as Figure 1 As shown, the content display system 10 includes a content display device 101, a display device 102, and an image acquisition device 103. The content display device 101 is connected to the display device 102 via a communication link, and the content display device 101 is also connected to the image acquisition device 103 via a communication link. This communication link can be a wired communication link or a wireless communication link; this disclosure does not limit the type of link.

[0074] It should be noted that the content display device 101, display device 102, and image acquisition device 103 in this disclosure can be one or more. For ease of understanding, Figure 1 Only one is shown in the image.

[0075] The display device 102 and the image acquisition device 103 can be located in at least one area within a preset scene. The display device 102 and the image acquisition device 103 can be located in the same area or in different areas.

[0076] The preset scenario is the application scenario of the technical solution provided in this disclosure. For example, the preset scenario can be public places such as shopping malls, scenic spots, exhibitions, and museums.

[0077] For example, such as Figure 2 As shown, the preset scene includes areas A, B, C, D, E, F, and G. Areas A, D, F, and G are equipped with a display device 102 and an image acquisition device 103, respectively. Areas B, C, and E are equipped only with the image acquisition device 103.

[0078] The content display device 101 is configured to acquire the gaze direction of a target user within the target range at the current time.

[0079] The target range can be a viewing range of the display device 102. The viewing range can be determined based on a best viewing distance, a visual angle, and other parameters of the display device 102.

[0080] For example, the viewing range can be a sector region in a circle with the center point of the display device 102 as the center and the best viewing distance as the radius, which overlaps with the visual angle.

[0081] For example, as shown in FIG. 3, the display device 102 is arranged on a corridor wall, and the target range 301 is the viewing range of the display device 102. A user located in the target range 301 can view the content displayed by the display device 102. Figure 3

[0082] It should be noted that when the display device 102 is a device with an audio playing function, the target range can be a broadcast range of the display device 102, which can be determined based on an output volume of the display device and geographical information of a current scene.

[0083] For example, in the broadcast range, a user receives an output volume of the display device 102 that is greater than a volume threshold. The broadcast range can be determined according to the output volume and a sound attenuation coefficient, or can be determined by field measurement.

[0084] It can be easily understood that when the display device 102 has both an image display function and an audio playing function, the target range can be determined based on one or more of a screen size, a best viewing distance, a visual angle, an output volume, and geographical information of a current scene of the display device 102.

[0085] For example, when the display device 102 has both an image display function and an audio playing function, the target range of the display device 102 can be a range overlapping the viewing range and the broadcast range.

[0086] In a possible implementation, the content display apparatus 101 can determine the gaze direction of the target user according to image information of a region in which the display device 102 is located, which is collected by the image collection apparatus 103.

[0087] The content display apparatus 101 is configured to determine the user portrait of the target user when the gaze direction of the target user at a current time is the display interface of the display device and a duration for which the target user gazes at the display interface exceeds a preset duration.

[0088] The preset duration can be set according to actual conditions, and the present disclosure does not limit this. For example, the preset duration can be 5 seconds.

[0089] ​The user portrait includes at least one of a first label and a second label; the first label is used to represent the attention point information of the target user, and the second label is used to represent the feature information of the target user. The first label can be determined based on the route trajectory of the target user, and the second label can be determined based on the image information of the target user. The image information includes at least one of face image information and body image information.

[0090] The attention point information of the target user refers to the thing that the target user currently pays attention to. For example, in a shopping mall, the thing that the target user currently pays attention to can include clothing, furniture, daily necessities, etc. In a scenic spot, the thing that the target user currently pays attention to can be a specific scenic spot, historical culture related to the scenic spot, etc.

[0091] The feature information of the target user can be the personal attribute of the target user, including at least one of the following: gender information, age information, emotion information, and clothing information of the target user. For example, the emotion information can be happy, sad, surprised, scared, angry, disgusted, etc. The clothing information can be the dressing style of the target user (for example, casual, fashionable, formal).

[0092] In a possible implementation, the content display device 101 is configured to determine the image information of the target user according to the image information collected by the image collection device 103.

[0093] For example, the body image information can be an image information including the complete body image of the target user or reflecting the feature of the body.

[0094] The content display device 101 is configured to display the target content corresponding to the user portrait on the display interface according to the user portrait.

[0095] In a possible implementation, the content display device 101 is configured to determine the target content according to the user portrait, and control the display device 102 to display the target content according to the content currently displayed on the display interface.

[0096] The display device 102 is configured to display the corresponding target content on the display interface in response to the control instruction of the content display device 101.

[0097] In a possible implementation, the display device 102 can perform a corresponding display operation based on the content currently displayed on the display interface.

[0098] For example, the display operation includes the following content:

[0099] When the content currently displayed on the display device 102 is default content, the display device 102 stops displaying the default content and displays the target content indicated by the content display device 101 in response to the control instruction of the content display device 101.

[0100] The default content can be content that the display device 102 plays by default when no control instruction of the content display apparatus 101 is received.

[0101] When the content currently displayed by the display device 102 is the target content indicated by the content display apparatus 101 at the previous time, the display device 102 displays the new target content indicated by the control instruction of the content display apparatus 101 at the current time after the target content indicated at the previous time is displayed in response to the control instruction of the content display apparatus 101 at the current time.

[0102] The control instruction is used to indicate the target content that the display device 102 needs to display.

[0103] It should be noted that during the period when the display device 102 displays the target content indicated at the previous time, if the content display apparatus 101 indicates multiple target contents to the display device 102 multiple times. The display device 102 can display the multiple target contents indicated by the content display apparatus 101 in time sequence, or can display the target content indicated by the content display apparatus 101 last.

[0104] The display device 102 is further configured to render the target content displayed in response to the special effect rendering instruction of the content display apparatus 101.

[0105] The special effect rendering instruction is used to instruct the display device 102 to display special effect information on the display interface.

[0106] For example, the special effect information is determined by the content display apparatus 101 according to the second label in the user portrait, and the special effect rendering instruction includes the special effect information displayed by the display device 102. The display device 102 is further configured to display the special effect information on the display interface according to the special effect information included in the special effect rendering instruction.

[0107] Alternatively, the display device 102 is configured with multiple special effect information. The display device 102 displays one or more special effects from the multiple special effect information on the display interface in response to the special effect rendering instruction.

[0108] The special effect information includes displaying an expression picture on the display interface, enlarging the display content of a specified area, outputting preset audio information, etc.

[0109] The display device 102 is further configured to stop displaying the current target content and display new content in response to the switching instruction of the content display apparatus 101.

[0110] The new content can be a new target content indicated by the content display apparatus 101, or a default content.

[0111] The image acquisition apparatus 103 is configured to acquire image information of the area where it is located.

[0112] For example, the image acquisition device 103 can acquire image information of the region according to a preset period, and send the acquired image information to the content display device 101. Correspondingly, the content display device 101 receives the image information sent by the image acquisition device 103.

[0113] The content display device, the display device and the image acquisition device in the embodiments of the present disclosure can have various forms.

[0114] As shown in the figure, the content display device 101, the display device 102 and the image acquisition device 103 are all independent entity devices. Figure 1

[0115] The image acquisition device 103 in the embodiments of the present disclosure is a device for converting image data into an analog signal or a digital signal through a photosensitive device, for example, the image acquisition device 103 includes a camera, a video camera, a camera. The image acquisition device 103 can also be a device with a camera function, for example, the image acquisition device 103 can be a mobile phone, a tablet computer, a notebook computer, a palm computer, a wearable device (such as a smart watch, a smart bracelet, a pedometer, etc.), a vehicle-mounted device, a flight device (such as a smart robot, a hot air balloon, a drone, an airplane), etc. with a camera function.

[0116] For example, the image acquisition device 103 in the embodiments of the present disclosure can also be an infrared imager or a night vision device, which is used to acquire image data of a dark region.

[0117] For example, the content display device 101 in the embodiments of the present disclosure can be a server, which includes:

[0118] The processor can be a general central processing unit (CPU), a microprocessor, an application-specific integrated circuit (ASIC), or one or more integrated circuits for controlling the execution of programs of the present disclosure.

[0119] The transceiver can be a device using any transceiver, which is used to communicate with other devices or communication networks, such as Ethernet, radio access network (RAN), wireless local area network (WLAN), etc.

[0120] ​The memory can be a read-only memory (ROM) or other type of static storage device that can store static information and instructions, a random access memory (RAM) or other type of dynamic storage device that can store information and instructions, an electrically erasable programmable read-only memory (EEPROM), a compact disc read-only memory (CD-ROM) or other optical disk storage, a magnetic disk storage or other magnetic storage devices, or any other medium capable of storing desired program code in the form of instructions or data structures and that can be accessed by a computer, but is not limited to this. The memory can exist independently and be connected to the processor through a communication line. The memory can also be integrated with the processor.

[0121] The display device 102 in the embodiments of the present disclosure is a device with an image display function. For example, the display device 102 can be a cathode ray tube (CRT) display, a liquid crystal display (LCD), or a light-emitting diode (LED) display.

[0122] It should be noted that the display device 102 in the embodiments of the present disclosure can also be a device with an audio playback function. At this time, the content involved in the present disclosure can be audio information. For example, the display device 102 can be a closed enclosure speaker, a bass-reflex enclosure speaker, an acoustic resistance enclosure speaker, a labyrinth enclosure speaker, and the like.

[0123] The display device 102 in the present disclosure can also be a device with both an image display function and an audio playback function, including an image display module and an audio playback module. At this time, the content involved in the present disclosure can be audio-visual information.

[0124] That is, the form of the content in the present disclosure is not limited. The content in the embodiments of the present disclosure can be static or dynamic image information, audio information, or multimedia information including text, sound, and images.

[0125] In another possible implementation, the display device 101, the display equipment 102, and the image acquisition device 103 can be coupled into the same device.

[0126] For example, Figure 4 This is a structural diagram of a display device 40 according to some embodiments. The display device 40 includes a content display module 401, a display module 402, and an image acquisition module 403. The content display module 401 is connected to the display module 402, and the content display module 401 is also connected to the image acquisition module 403.

[0127] Figure 4 The content display module 401 in the middle is... Figure 1 Another product form of the content display device 101. Figure 4 The display module 402 in the middle is Figure 1 Another product form of the display device 102. Figure 4 The image acquisition module 403 in the middle is Figure 1 Another product form of the image acquisition device 103.

[0128] It should be noted that the various embodiments of this disclosure can be referenced or learned from each other. For example, the same or similar steps, method embodiments, system embodiments and device embodiments can be referenced from each other without limitation.

[0129] Figure 5 This is a flowchart illustrating a content display method according to some embodiments. Figure 5 As shown, the method includes the following steps:

[0130] Step 501: The content display device acquires the gaze direction of the target user within the target range at the current time.

[0131] The target range can be determined based on the content format output by the display device and the product parameters of the display device. For details, please refer to the relevant descriptions above, which will not be repeated here.

[0132] In one possible implementation, the content display device can determine the target user's gaze direction based on image information within the area where the display device is located. This image information is acquired by an image acquisition device within that area.

[0133] For example, the content display device determines the target user's location information and the target user's facial image information based on image information within the area where the display device is located.

[0134] For example, the content display apparatus inputs the image information into the user recognition model to obtain the body image information and the face image information of the target user. The content display apparatus determines the position relationship of the target user according to the position of the body image information in the image information.

[0135] When it is determined that the target user is in the target range, the content display apparatus determines the gaze direction of the target user according to the face image information of the target user.

[0136] For example, Figure 6 The image acquisition apparatus 602 in the display device 601 acquires image information of the area in which the display device 601 is located, and sends the image information to the content display apparatus. The content display apparatus determines the position information of the target user 603 as the area 605 according to the image information. When it is determined that the target user 603 is in the target range 604, the content display apparatus can obtain the face image information of the target user 603 from the image information, and determine the gaze direction of the target user 603 according to the face image information.

[0137] The content display apparatus can perform face orientation recognition on the face image information to determine the key point information (for example, the position point information of organs such as eyebrows, eyes, mouth, and nose) of the target user 603 in the face image information, and determine the gaze direction of the target user 603 according to the key point information of the target user 603. The gaze direction of the target user 603 can be vector coordinate information, which is used to represent the included angle between the image acquisition apparatus 602 and the gaze direction of the target user 603.

[0138] For example, the content display apparatus can perform the above operation through a face orientation recognition model: the content display apparatus inputs the face image information into the face orientation recognition model to obtain the vector coordinate information of the target user.

[0139] The face orientation recognition model can be trained by a plurality of sets of face image information and corresponding vector coordinate information representing the gaze direction. The face orientation recognition model can be a neural network (NNs) model.

[0140] It should be noted that when the image acquisition apparatus is an image acquisition module in the display device, the content display apparatus obtains the gaze direction of the target user in the target range at the current time in a manner similar to the above-mentioned solution, which will not be described here.

[0141] Step 502, the content display apparatus determines whether the duration for which the target user gazes at the display interface of the display device exceeds a preset duration.

[0142] In a possible implementation, the content display apparatus can decompose the angle between the image acquisition apparatus and the gaze direction of the target user into a longitudinal angle and a transverse angle based on the display interface of the display device as a section plane, and determine whether the target user gazes at the display interface according to the longitudinal angle and the transverse angle.

[0143] In an example, the content display apparatus can determine whether the target user gazes at the display interface according to whether the longitudinal angle is located in a longitudinal display angle interval and whether the transverse angle is located in a transverse display angle interval.

[0144] The longitudinal display angle interval refers to an angle interval formed by the height of the display device and the eye position of the target user. For example, as shown in FIG. 7A, the longitudinal angle interval can be an angle A to an angle C in FIG. 7B. Figure 7A Figure 7A The transverse display angle interval refers to an angle interval formed by the width of the display device and the eye position of the target user. For example, as shown in FIG. 7C, the transverse angle interval can be an angle D to an angle F in FIG. 7D. Figure 7B Figure 7B The following describes a process in which the content display apparatus determines whether the target user gazes at the display interface, in combination with FIG. 7A and FIG. 7C. Figure 7A Figure 7B For example, as shown in FIG. 7A,

[0145] For example, as shown in FIG. 7A, Figure 7A Figure 7A is a side view of a content display scene provided according to some embodiments. The thick line represents the height information of the display device 701.

[0146] The content display apparatus determines the longitudinal display angle interval of the display device 701 according to the position information and the height information of the display device 701, the position information of the image acquisition apparatus 702, and the position information of the target user 703, that is, the angle A to the angle C in FIG. 7B. Figure 7A

[0147] The content display apparatus can further determine the longitudinal angle B between the image acquisition apparatus 702 and the gaze direction of the target user 703 according to the face image information of the target user 703.

[0148] When the longitudinal angle B is located in the longitudinal display angle interval, it indicates that the gaze direction of the target user 703 satisfies the condition of gazing at the display interface in the longitudinal direction, and the content display apparatus further needs to determine whether the target user 703 satisfies the condition of gazing at the display interface in the transverse direction.

[0149] For example, as shown in FIG. 7C, Figure 7B Figure 7B is a top view of a content display scene provided according to some embodiments. The thick line represents the width information of the display device 701. ​​​​​​

[0150] The content display device determines the horizontal display angle interval of the display device 701 according to the position information of the display device 701, the width information, the position information of the image collection device 702, and the position information of the target user 703, that is, Figure 7B the included angle D and the included angle F in FIG. 6.

[0151] The content display device can further determine the horizontal included angle E between the image collection device 702 and the gaze direction of the target user 703 according to the face image information of the target user 703.

[0152] When the horizontal included angle E is located in the horizontal display angle interval, it indicates that the gaze direction of the target user 703 satisfies the condition of gazing at the display interface in the horizontal direction.

[0153] When both the horizontal included angle and the vertical included angle of the target user 703 satisfy the condition of gazing at the display interface, the content display device determines that the gaze direction of the target user 703 at the current time is the display interface of the display device 701.

[0154] Conversely, when at least one of the horizontal included angle and the vertical included angle of the target user 703 does not satisfy the condition of gazing at the display interface, the content display device determines that the gaze direction of the target user 703 at the current time is not the display interface of the display device 701.

[0155] It should be noted that the content display device can continuously determine whether the gaze direction of the target user 703 at the current time is the display interface of the display device 701 within a preset time length, and further determine whether the time length of the target user 703 gazing at the display interface of the display device 701 exceeds the preset time length. The preset time length can be set according to actual conditions, and the present disclosure does not limit this.

[0156] When the gaze direction of the target user 703 at the current time is the display interface of the display device 701 and the time length of the target user 703 gazing at the display interface exceeds the preset time length, the content display device can perform the following step 503.

[0157] Step 503, the content display device determines a user portrait of the target user.

[0158] The user portrait includes at least one of a first label and a second label; the first label is used to represent the attention point information of the target user, and the second label is used to represent the feature information of the target user.

[0159] The attention point information of the target user refers to the things that the target user currently pays attention to. For example, in a shopping mall, the things that the target user currently pays attention to can include clothing, furniture, daily necessities, etc. In a scenic spot, the things that the target user currently pays attention to can be specific scenic spots, historical and cultural information related to the scenic spots, etc.

[0160] The characteristic information of the target user can be personal attributes of the target user, including at least one of the following: gender information, age information, emotion information, and clothing information of the target user. For example, the emotion information can be happy, sad, surprised, scared, angry, disgusted, and the like. The clothing information can be the dressing style of the target user (e.g., casual, fashionable, formal).

[0161] The first label in the user portrait can be determined according to a route trajectory of the target user. The route trajectory includes one or more regions passed by the target user and time information of the target user reaching the one or more regions.

[0162] The second label in the user portrait can be determined according to image information of the target user. The image information of the target user includes at least one of face image information and body image information.

[0163] In a possible implementation, the content display apparatus can obtain the image information of the target user, and input the image information of the target user into a characteristic information recognition model to obtain the second label of the target user.

[0164] The characteristic information recognition model can be used to determine the second label of the user. For example, the characteristic information recognition model can be a neural network model. The characteristic information recognition model can include one model or multiple sub-models. When the characteristic information recognition model is one model, the content display apparatus can input the image information into the characteristic information recognition model to obtain a plurality of second labels corresponding thereto.

[0165] When the characteristic information recognition model is multiple sub-models, the multiple sub-models correspond to the multiple second labels one by one. The content display apparatus can input the image information into each sub-model respectively to obtain a plurality of second labels corresponding thereto.

[0166] In a possible implementation, there are multiple target users in the target range, the first target label of the target user is a label that is the same and has the largest quantity in the first labels of the multiple target users, and the second target label of the target user is a label that is the same and has the largest quantity in the second labels of the multiple target users. The user portrait of the target user includes at least one of the first target label and the second target label.

[0167] In this way, for each target user in the multiple target users, the content display apparatus determines the user portrait of the target user from the whole according to the label of the target user, so as to subsequently control the display device to display the corresponding target content based on the user portrait, so as to meet the needs of the multiple target users.

[0168] In step 504, the content display apparatus displays target content corresponding to the user portrait on the display interface of the display device according to the user portrait. In step 504, the content display apparatus displays target content corresponding to the user portrait on the display interface of the display device according to the user portrait.

[0169] The target content is determined by a user portrait of the target user. For example, in a shopping mall, the target content can be an introduction content of a product that interests the target user. In a museum, the target content can be an explanation content of an exhibit that interests the target user.

[0170] In a possible implementation, the content display apparatus can determine the target content corresponding to the user portrait according to the user portrait, and display the target content on the display interface of the display device. Specifically, reference can be made to the description below. Figure 8

[0171] Based on the above technical solution, the content display apparatus in the present disclosure can obtain the gaze direction of the target user in the target range at the current time, and when the duration that the target user gazes at the display interface of the display device exceeds the preset duration, further determine the user portrait of the target user, so as to display the target content corresponding to the user portrait on the display interface according to the user portrait. Since the user portrait includes at least one of the first label for representing the attention point information of the target user and the second label for representing the characteristic information of the target user, the content display apparatus in the present disclosure can display the corresponding target content for the target user based on two dimensions of the attention content and the personal attribute of the target user, meet the actual needs of the user, and improve the promotion effect.

[0172] In the following, the process that the content display apparatus displays the target content corresponding to the user portrait on the display interface according to the user portrait is introduced.

[0173] As a possible embodiment of the present disclosure, in combination with Figure 5 as shown in Figure 8 , the above step 504 can be implemented by the following steps 801-802.

[0174] Step 801, the content display apparatus determines the target content according to the user portrait.

[0175] It should be noted that since the user portrait can represent the attention point information and the characteristic information of the target user, the target content determined based on the user portrait can better meet the actual needs of the target user.

[0176] In a possible implementation, the content display apparatus can determine the target content corresponding to the user portrait from the preset content set through a classification algorithm.

[0177] For example, the classification algorithm can be a K-nearest neighbor (KNN) algorithm, a decision tree algorithm, or a Bayesian classification algorithm.

[0178] ​The preset content set can be set according to an actual application scenario. For example, in a shopping mall, the preset content set can be the introduction content of the goods sold by each store in the shopping mall. In a museum, the preset content set can be the explanation content of each exhibit.

[0179] In another possible implementation, the content display apparatus determines the target content according to the user portrait of the target user and a third label of the display device.

[0180] The third label is used to represent the device information of the display device. For example, the device information can be a display content set preconfigured for the display device. The display content set can be determined according to the area information of the area where the display device is located. In this way, when determining the target content, the content display apparatus can be based on both the user portrait factors of the target user and the device factors of the display device, so that the determined target content is more in line with the actual needs of the user.

[0181] In step 802, the content display apparatus controls the display device to display the target content according to the content currently displayed by the display interface.

[0182] The content currently displayed by the display interface can be the first content or the second content. The first content is related to the user portrait of the user. The second content is not related to the user portrait of the user. For example, the first content refers to the target content determined by the content display apparatus at the previous time. The second content refers to default content, no picture display, and the like.

[0183] In a possible implementation, when the content displayed by the display interface is the first content, after the display of the first content is completed, the content display apparatus can control the display device to display the target content. When the content displayed by the display interface is the second content, the content display apparatus can control the display device to stop displaying the second content and display the target content. For example, when the content currently displayed by the display interface is the target content determined by the content display apparatus at the previous time (i.e., the first content), the display device can display the new target content determined by the content display apparatus after the display of the content is completed, thereby avoiding the previous user's viewing experience from being affected.

[0184] When the content currently displayed by the display interface is default content, or the display device is in a standby state and the display interface is a black screen (i.e., the second content), the display device can directly display the target content determined by the content display apparatus.

[0185] It should be noted that during the period when the display interface displays the first content, the content display apparatus can determine multiple target contents. The display device can display the multiple target contents determined by the content display apparatus in time sequence, or can display the latest target content determined from the multiple target contents determined by the content display apparatus.

[0186] Based on the above technical solution, the content display device in the present disclosure can determine the target content according to the user portrait, and control the display device to display the target content according to the content currently displayed on the display interface. When the content displayed on the display interface is the first content related to the user portrait, the content display device controls the display device to display the target content after the display of the first content is completed. When the content displayed on the display interface is the second content unrelated to the user portrait, the content display device can control the display device to stop displaying the second content and display the target content. In this way, the above technical solution of the present disclosure can not only guarantee the actual needs of the target user, but also avoid affecting the viewing experience of the previous user.

[0187] Next, the process of determining the target content by the content display device according to the user portrait is introduced.

[0188] As a possible embodiment of the present disclosure, in combination with Figure 8 As shown in Figure 9 The step 801 can also be implemented by the following steps 901-903.

[0189] Step 901: For each label in the user portrait of the target user, the content display device determines the weight of each content corresponding to the label in the preset content set.

[0190] Each label in the user portrait is a first label and / or at least one second label.

[0191] The preset content set can be set according to the actual application scenario. For example, in a shopping mall, the preset content set can be the introduction content of the goods sold by each store in the shopping mall. In a museum, the preset content set can be the explanation content of each exhibit.

[0192] It is easy to understand that the weight of the content in the preset content set corresponding to the label in the user portrait can represent the degree of interest of the user with the label to the content. The weight of each content can be set according to the user portrait and the interested content of the user in the historical data.

[0193] For example, the user portrait includes a first label A1, a second label B1 and a second label B2. The preset content set includes content C1, content C2 and content C3.

[0194] For the first label A1, the content display device determines that the weight of the corresponding content C1 is 1, the weight of the content C2 is 2, and the weight of the content C3 is 3.

[0195] For the second label B1, the content display device determines that the weight of the corresponding content C1 is 0.5, the weight of the content C2 is 2, and the weight of the content C3 is 0.5.

[0196] For the second label B2, the content display device determines that the weight of the corresponding content C1 is 1.5, the weight of content C2 is 1, and the weight of content C3 is 1.

[0197] It should be noted that the weight of each content in the preset content set corresponding to each tag can be set according to the actual situation, and this disclosure does not limit it.

[0198] Step 902: For each piece of content in the preset content set, the content display device calculates the sum of the weights of each tag corresponding to that content.

[0199] Based on the example in step 901 above, for content C1, the content display device calculates that the sum of the weights corresponding to content C1 is 3.

[0200] For content C2, the content display device calculates that the sum of the weights corresponding to content C2 is 5.

[0201] For content C3, the content display device calculates that the sum of the weights corresponding to content C3 is 4.5.

[0202] Step 903: The content display device selects the content with the largest sum of weights in the preset content set as the target content.

[0203] It should be noted that the higher the sum of the weights, the greater the target user's interest in the content. Based on this, the content display device selects the content with the highest sum of weights as the target content, thus meeting the actual needs of the target user.

[0204] It should be noted that the technical solutions described in steps 901-903 above also apply to the content display device determining the target content based on the user profile and the third tag of the display device. Further details will not be elaborated here.

[0205] Based on the above technical solution, the content display device in this disclosure can determine the target content based on the weight of each tag in the user profile relative to each content in the preset content set, thereby improving the accuracy of the content display device in identifying the content of interest to the target user and meeting the actual needs of the target user.

[0206] As another possible embodiment of this disclosure, combined with Figure 8 ,like Figure 10 As shown, step 801 above can also be achieved through the following steps 1001-1002.

[0207] Step 1001: The content display device inputs each tag from the target user's user profile into a preset decision model to obtain multiple predicted contents.

[0208] The preset decision model includes a plurality of decision sub-models, each of which is configured to determine a predicted content according to each label. Each decision sub-model includes a plurality of decision layers, each of which includes at least one decision node. The plurality of decision layers correspond to the plurality of labels in the user portrait one by one. Each decision node included in the same decision layer is configured to make a decision for the label corresponding to the decision layer.

[0209] Exemplarily, Figure 11 A structural diagram of a decision sub-model according to some embodiments is shown. The decision sub-model includes three decision layers, the decision layer 1 includes one decision node, the decision layer 2 includes two decision nodes, and the decision layer 3 includes four decision nodes. The content display device makes a decision for the user portrait of the target user according to the decision sub-model, including the following three steps:

[0210] Step 1: The content display device performs a decision operation of the decision node in the decision layer 1.

[0211] The content display device performs a decision operation of the decision node 1, and determines whether the focus point information of the target user satisfies the judgment condition of the decision node 1.

[0212] In the case that the focus point information of the target user is the focus point A, the content display device continues to perform a decision operation of the decision node 2.

[0213] In the case that the focus point information of the target user is not the focus point A, the content display device continues to perform a decision operation of the decision node 3.

[0214] Step 2: The content display device performs a decision operation of the corresponding decision node in the decision layer 2 according to the judgment result of the decision layer 1.

[0215] For example, when the content display device performs a decision operation of the decision node 2, the content display device determines whether the gender of the target user satisfies the judgment condition of the decision node 2.

[0216] In the case that the gender of the target user is male, the content display device continues to perform a decision operation of the decision node 4.

[0217] In the case that the gender of the target user is female, the content display device continues to perform a decision operation of the decision node 5.

[0218] When the content display device performs a decision operation of the decision node 3, the execution operation of the content display device can refer to the decision operation of the decision node 2 and Figure 11 which will not be repeated here.

[0219] Step 3: The content display device performs a decision operation of the corresponding decision node in the decision layer 3 according to the judgment result of the decision layer 2.

[0220] For example, when the content display device performs the decision operation of the decision node 7, the content display device determines whether the age of the target user satisfies the determination condition of the decision node 7.

[0221] When the age of the target user is less than or equal to 15 years old, the content display device determines that the predicted content is the content 7. When the age of the target user is greater than 15 years old and less than or equal to 25 years old, the content display device determines that the predicted content is the content 8. When the age of the target user is greater than 25 years old, the content display device determines that the predicted content is the content 9.

[0222] When the content display device performs the decision operation of other decision nodes in the decision layer 3, the execution operation of the content display device can refer to the decision operation of the decision node 7 described above and the technical solutions of the decision node 8, which will not be described herein again. Figure 11

[0223] Step 1002, the content display device determines the same and the most number of predicted contents in the plurality of predicted contents as the target content.

[0224] Based on the above technical solutions, the content display device in the present disclosure can determine the target content through a plurality of decision sub-models, thereby avoiding the problem of low prediction accuracy due to too few training samples, or deviation from the actual situation due to overfitting of too many training samples. Therefore, the above embodiments of the present disclosure can improve the accuracy of the content display device in identifying the content of interest of the target user, and meet the actual needs of the target user.

[0225] Next, the process in which the content display device determines the first label in the user portrait will be described.

[0226] As a possible embodiment of the present disclosure, in combination with Figure 5 As shown in Figure 12 In the case where the user portrait includes the first label, the above step 503 can also be implemented through the following steps 1201-1203.

[0227] Step 1201, the content display device acquires the route trajectory of the target user.

[0228] The route trajectory includes one or more regions passed by the target user in chronological order, and time information of the target user reaching the one or more regions.

[0229] In a possible implementation manner, the content display device can acquire the face image information of the target user, and determine the identification information of the target user according to the face image information of the target user and the first preset database. Then, the content display device can acquire the route trajectory of the target user from the second preset database according to the identification information of the target user. ​

[0230] The first preset database includes a plurality of face image information and identification information corresponding to each face image information, and the second preset database includes identification information of a plurality of users and a route trajectory corresponding to each user.

[0231] It should be noted that the first preset database and the second preset database can be the same database, which simultaneously stores identification information of a plurality of users, face image information corresponding to each identification information, and a route trajectory corresponding to each identification information.

[0232] For example, the content display device can input the face image information of the target user into the face recognition model to obtain the face feature of the target user. The content display device can also input the face image information in the first preset database into the face recognition model to obtain the face feature corresponding to the face image information in the first preset database.

[0233] Then, the content display device can match the face feature of the target user with each face feature corresponding to the face image information in the first preset database, and take the identification information corresponding to the face image information of the target user in the first preset database as the identification information of the target user.

[0234] It should be noted that in the case that there is no face image information in the first preset database that matches the target user, the content display device can create new identification information in the first preset database, and take the face image information of the target user as the face image information corresponding to the identification information.

[0235] Alternatively, in the case that there is no face image information in the first preset database that matches the target user, the content display device can reacquire a new target user and perform the above steps 501-504 to display new target content on the display device.

[0236] The face image information of the target user can be an image with a quality score greater than a preset quality threshold, thereby avoiding the problem of matching errors caused by poor image quality of the face image information. The higher the quality score, the clearer the face image information and the higher the recognition rate.

[0237] For example, the content display device can input the face image information of the target user into the image quality recognition model to obtain the quality score of the face image information.

[0238] The face recognition model and the image quality recognition model described above can be a neural network model.

[0239] Step 1202, the content display device determines a target residence area of the target user according to the route trajectory.

[0240] The target residence area is one or more areas in which the target user stays for a duration longer than a preset duration. The preset duration can be set according to actual conditions, and the present disclosure does not limit this.

[0241] In a possible implementation, the content display device determines a target duration and a target distance of the target user in the route track.

[0242] When the target duration and the target distance meet a first preset condition, the content display device determines that the target residence area of the target user is the first area.

[0243] The target duration is a time difference between a time at which the target user arrives at the first area and a time at which the target user arrives at the second area, and the target distance is a distance between the first area and the second area. The first area is any area passed through by the target user, and the second area is a next area passed through by the target user after the first area.

[0244] The distance between the first area and the second area can be determined according to map information of a preset scene. The map information of the preset scene includes position information of each area and paths between the areas.

[0245] For example, the content display device can determine the distance between the first area and the second area based on the map information of the preset scene by using a path planning algorithm. The path planning algorithm can be a breadth first search (BFS) algorithm, a Dijkstra algorithm, or an A star algorithm.

[0246] The first preset condition can include that a difference between a predicted distance of the target user and a preset distance threshold is greater than the target distance. The predicted distance can be determined according to the target duration and an average speed of the target user. The average speed can be determined according to speeds of users in the preset scene.

[0247] It should be noted that the time difference between the time at which the target user arrives at the first area and the time at which the target user arrives at the second area includes a duration from the time at which the target user arrives at the first area to the time at which the target user departs from the first area, and a duration from the time at which the target user departs from the first area to the time at which the target user arrives at the second area. That is, the longer the target duration of the target user is, the longer the predicted distance determined is. When the difference between the predicted distance of the target user and the preset distance threshold is greater than the target distance, it indicates that the duration from the time at which the target user arrives at the first area to the time at which the target user departs from the first area, that is, the duration in which the target user stays in the first area, exceeds the preset duration.

[0248] Therefore, the content display device can determine whether the first area is the target residence area by determining whether the target duration and the target distance meet the first preset condition.

[0249] It should be noted that, since the moving speed of different users will have certain differences, the above-mentioned preset distance threshold value can be positive or negative. For example, the content display device can determine the preset distance threshold value based on the characteristic information of the target user such as age, gender, etc., so as to eliminate the difference in moving speed of different users and avoid affecting the determination of the target residence area of the target user in the present disclosure.

[0250] For example, the first preset condition can be represented by the following formula 1:

[0251] t*v-d>D Formula 1

[0252] Wherein, t represents the target duration of the target user, v represents the average speed, d represents the preset distance threshold value, and D represents the target distance.

[0253] Step 1203, the content display device determines the first label according to the area information of the target residence area.

[0254] Wherein, the area information is used to represent the unique attributes of the area.

[0255] For example, in a shopping mall, the area information can be the item information (such as clothing, furniture, etc.) sold by the stores in the area. In a scenic area, the area information can be the scenic spot information of the scenic spots in the area, related history and culture, etc.

[0256] In the case where the target residence area is at least one area, the content display device can take the area information of the target residence area where the target user last arrived in the at least one area as the first label of the target user.

[0257] Alternatively, the content display device can take the area information with the same area information and the largest number in the at least one area as the first label of the target user.

[0258] In the case where the content display device determines that the target user does not have a target residence area, the first label can be set to empty.

[0259] Based on the above technical solution, the content display device in the present disclosure can determine the target residence area of the target user according to the route trajectory of the target user, and then determine the first label of the target user according to the area information of the target residence area. Since the area where the target user stays for a long time is usually the area that the target user pays attention to, the first label determined based on the above technical solution can accurately represent the attention point information of the target user, which is beneficial to subsequent determination of target content based on the first label to meet the actual needs of users.

[0260] Next, the process of generating the route trajectory of the user by the content display device will be introduced.

[0261] As a possible embodiment of the present disclosure, as shown in Figure 13 The content display device can generate the route track of the user in the following steps 1301-1302.

[0262] In step 1301, when detecting that the first user arrives at the third area, the content display device determines time information of the first user arriving at the third area.

[0263] The third area is any area in the target scene, and the first user is any user in the preset scene.

[0264] In a possible implementation, the content display device can acquire image information of the area where each image acquisition device is located according to a preset period, and detect whether the first user exists according to the acquired image information.

[0265] In the case where the first user exists in the acquired image information, the content display device regards the area corresponding to the image information as the area arrived by the first user, and regards the time when the image acquisition device acquires the image information as the time when the first user arrives at the area.

[0266] In step 1302, the content display device generates the route track of the first user according to the time information of the first user arriving at the third area.

[0267] In a possible implementation, the content display device can determine the identification information of the first user according to the face image information of the first user and a first preset database. Then, the content display device can store the third area and the time information of the first user arriving at the third area in a second preset database according to the identification information of the first user, and establish a mapping relationship with the identification information of the first user, thereby generating the route track of the first user.

[0268] When the first user passes through other areas before passing through the third area, based on the above method, the route track of the first user includes the third area, the other areas, the time information of the first user arriving at the third area, and the time information of the first user arriving at the other areas.

[0269] The data included in the route track can be sorted in chronological order.

[0270] It should be noted that the process of determining the identification information of the first user by the content display device can refer to the related description in step 1201.

[0271] In a case where the face image information matching the first user does not exist in the first preset database, the content display device can create new identification information in the first preset database, and store the face image information of the first user as the face image information corresponding to the identification information. The content display device can also create the same identification information in the second preset database, and store the third region and the time information of the first user reaching the third region in the second preset database, to establish a mapping relationship with the identification information of the first user.

[0272] Based on the above technical solutions, the content display device in the present disclosure can determine the time information of the user reaching each region according to the region images of each region collected by the image collection device, thereby generating the route trajectory of the user, so as to facilitate subsequent determination of the first label based on the route trajectory of the user.

[0273] As a possible embodiment of the present disclosure, in combination with Figure 5 As shown in Figure 14 The content display method further includes the following step 1401.

[0274] In step 1401, the content display device determines the special effect information to be displayed by the display device according to the second label in the user portrait.

[0275] The special effect information is used to render the target content.

[0276] For example, the content display device can control the display device to display an expression picture on the display interface, enlarge the display content of the specified region, output preset audio information, and the like.

[0277] It should be noted that the content display device can render the target content before the display device displays the target content, or render the subsequent part of the target content in the process of the display device displaying the target content, so as to achieve the effect of instant rendering of special effects.

[0278] Based on the above technical solutions, the content display device in the present disclosure can determine the special effect information according to the second label of the target user, and render the target content, thereby attracting the attention of the target user and improving the promotion effect.

[0279] As a possible embodiment of the present disclosure, in combination with Figure 5 As shown in Figure 15 The content display method further includes the following steps 1501-1503.

[0280] In step 1501, the content display device acquires the body image information of the target user.

[0281] In a possible implementation, the content display apparatus can acquire image information of a region where the display device is located collected by the image collection apparatus, and determine user portrait information of the target user according to the image information.

[0282] In step 1502, the content display apparatus performs image recognition on the human body image information of the target user, to obtain user gesture information of the target user.

[0283] The human body image information of the target user can be an image frame, and in this case, the user gesture information of the target user obtained by the content display apparatus is a static gesture, for example, a thumbs-up gesture.

[0284] The human body image information of the target user can also be multiple image frames, and in this case, the user gesture information of the target user obtained by the content display apparatus is a dynamic gesture, for example, a waving gesture.

[0285] In a possible implementation, the content display apparatus can input the human body image information of the target user into a gesture recognition model, to obtain the user gesture information of the target user. For example, the gesture recognition model can be a neural network model.

[0286] When the user gesture information of the target user is preset gesture information, the content display apparatus performs the following step 1503.

[0287] In step 1503, the content display apparatus controls the display device to stop displaying the target content, and displays third content.

[0288] The third content can be new target content determined by the content display apparatus according to user portrait information of the target user at the next moment, or can be default content.

[0289] Based on the above technical solutions, the content display apparatus in the present disclosure can switch the content displayed by the display device in response to the gesture of the current target user, so that the user can independently select the content to watch according to his / her own preferences, thereby better meeting the actual needs of the user.

[0290] As a possible embodiment of the present disclosure, in combination with Figure 15 As shown in Figure 16 When there are multiple target users in the target range, the above steps 1501-1502 can be replaced by the following steps 1601-1602.

[0291] In step 1601, the content display apparatus acquires human body image information of multiple target users.

[0292] In one possible implementation, the content display device can acquire image information of the area where the display device is located, captured by the image acquisition device, and determine multiple user profiles based on the image information. Each user profile corresponds one-to-one with a target user.

[0293] Step 1602: The content display device performs image recognition on the human body image information of the first target user to obtain the user gesture information of the first target user.

[0294] The first target user is the target user among multiple target users who meets the second preset condition. The second preset condition includes: the hand image area included in the user profile information is the largest, or the target user's position is the closest to the display device.

[0295] When the user gesture information of the first target user is the preset gesture information, the content display device executes the above step 1503.

[0296] Based on the above technical solution, the content display device in this disclosure can switch the content displayed on the display device according to the user gesture information of the first target user that meets the second preset condition. Since the first target user is the user who is closest to the display device or whose corresponding hand image area is the largest among multiple target users, the content display device can more accurately identify the user's gesture information, thereby avoiding the problem of misoperation caused by gesture recognition error and improving the user's viewing experience.

[0297] This disclosure embodiment can divide the content display device into functional modules or functional units according to the above method examples. For example, each function can be divided into its own functional modules or functional units, or two or more functions can be integrated into one processing module. The integrated module can be implemented in hardware or as a software functional module or functional unit. The module or unit division in this disclosure embodiment is illustrative and represents only one logical functional division; in actual implementation, other division methods may be used.

[0298] like Figure 17 The diagram shown is a structural diagram of a content display device 170 provided according to some embodiments. The device includes a processing unit 1701 and an acquisition unit 1702.

[0299] The acquisition unit 1702 is configured to acquire the gaze direction of the target user within the target range at the current time.

[0300] The processing unit 1701 is configured to determine a user portrait of the target user when the gaze direction of the target user at the current time is the display interface of the display device and the duration for which the target user gazes at the display interface exceeds a preset duration. The user portrait includes at least one of a first label and a second label. The first label is used to represent the attention point information of the target user, and the second label is used to represent the feature information of the target user.

[0301] The processing unit 1701 is further configured to display target content corresponding to the user portrait on the display interface according to the user portrait.

[0302] In some embodiments, the processing unit 1701 is configured to determine target content according to the user portrait, and control the display device to display the target content according to the content currently displayed on the display interface.

[0303] In some embodiments, the processing unit 1701 is configured to determine, for each label in the user portrait of the target user, a weight of each content in a preset content set corresponding to the label, calculate, for each content in the preset content set, a sum of weights of the content corresponding to each label, and take the content in the preset content set with the largest sum of weights as the target content.

[0304] In some embodiments, the processing unit 1701 is configured to input each label in the user portrait of the target user into a preset decision model to obtain a plurality of predicted contents. The preset decision model includes a plurality of decision sub-models. Each decision sub-model is used to determine a predicted content according to each label. The same and most numerous predicted content in the plurality of predicted contents is taken as the target content.

[0305] In some embodiments, the processing unit 1701 is configured to control the display device to display the target content after the display of the first content is completed when the content displayed on the display interface is the first content. The first content is related to the user portrait of the user. When the content displayed on the display interface is the second content, the processing unit 1701 is configured to control the display device to stop displaying the second content and display the target content. The second content is not related to the user portrait of the user.

[0306] In some embodiments, the acquisition unit 1702 is configured to acquire a route trajectory of the target user. The route trajectory includes one or more regions passed by the target user in a time sequence and time information of the target user reaching the one or more regions. The processing unit 1701 is configured to determine a target residence region of the target user according to the route trajectory. The target residence region is a region in the one or more regions in which the target user stays for a duration exceeding a preset duration. The processing unit 1701 is further configured to determine the first label according to region information of the target residence region.

[0307] In some embodiments, the acquisition unit 1702 is configured to acquire the face image information of the target user; the processing unit 1701 is configured to determine the identification information of the target user according to the face image information of the target user and a first preset database; the first preset database includes a plurality of face image information and identification information corresponding to each face image information; the processing unit 1701 is further configured to acquire the route trajectory of the target user from a second preset database according to the identification information of the target user; the second preset database includes a plurality of identification information of users and a route trajectory of each user.

[0308] In some embodiments, the processing unit 1701 is configured to determine a target duration and a target distance of the target user in the route trajectory; the target duration is a time difference between a time when the target user arrives at a first area and a time when the target user arrives at a second area; the target distance is a distance between the first area and the second area; the first area is any area passed by the target user, and the second area is a next area after the target user passes the first area; in a case where the target duration and the target distance satisfy a first preset condition, the target user is determined to have a target residence area as the first area.

[0309] In some embodiments, the acquisition unit 1702 is configured to acquire image information of the target user; the image information of the target user includes at least one of face image information and body image information; the processing unit 1701 is configured to input the image information of the target user into a feature information recognition model to obtain a second label of the target user.

[0310] In some embodiments, there are a plurality of target users in the target range, a first target label of the target user is a label that is the same and has the largest quantity in first labels of the plurality of target users, and a second target label of the target user is a label that is the same and has the largest quantity in second labels of the plurality of target users; the user portrait of the target user includes at least one of the first target label and the second target label.

[0311] In some embodiments, the processing unit 1701 is configured to determine special effect information to be displayed by the display device according to the second label in the user portrait; the special effect information is used to render target content.

[0312] In some embodiments, the acquisition unit 1702 is configured to acquire body image information of the target user; the processing unit 1701 is configured to perform image recognition on the body image information of the target user to obtain user gesture information of the target user; the processing unit 1701 is further configured to control the display device to stop displaying the target content and display third content when the user gesture information of the target user is preset gesture information.

[0313] When implemented by hardware, the acquisition unit 1702 in the embodiments of the present disclosure can be integrated on a communication interface, and the processing unit 1701 can be integrated on a processor. The specific implementation manner is as shown in Figure 18

[0314] Figure 18 Another possible structural schematic diagram of the content display device involved in the above embodiments is shown. The content display device 180 includes a processor 1802 and a communication interface 1803. The processor 1802 is configured to control and manage the actions of the content display device 180, for example, to perform the steps performed by the processing unit 1701 described above, and / or to perform other processes of the technologies described herein. The communication interface 1803 is configured to support the communication of the content display device 180 with other network entities, for example, to perform the steps performed by the acquisition unit 1702 described above. The content display device 180 can further include a memory 1801 and a bus 1804, and the memory 1801 is configured to store the program code and data of the content display device 180.

[0315] The memory 1801 can be a memory or the like in the content display device 180, which can include a volatile memory such as a random access memory, and can also include a non-volatile memory such as a read-only memory, a flash memory, a hard disk or a solid state disk, and can also include a combination of the above-mentioned kinds of memories.

[0316] The processor 1802 described above can be various exemplary logical blocks, modules and circuits described in combination with the above content of the present disclosure. The processor can be a central processing unit, a general purpose processor, a digital signal processor, an application specific integrated circuit, a field programmable gate array or other programmable logic device, a transistor logic device, a hardware component or any combination thereof. It can implement or execute various exemplary logical blocks, modules and circuits described in combination with the above content of the present disclosure. The processor can also be a combination of computing functions, such as one or more microprocessor combinations, combinations of DSP and microprocessor, etc.

[0317] The bus 1804 can be an extended industry standard architecture (EISA) bus or the like. The bus 1804 can be divided into an address bus, a data bus, a control bus, etc. For ease of representation, Figure 18 only one thick line is used in the figure, but it does not mean that there is only one bus or only one type of bus.

[0318] Figure 18 The content display device 180 in the above embodiments can also be a chip. The chip includes one or more than two (including two) processors 1802 and communication interfaces 1803.​

[0319] In some embodiments, the chip further includes a memory 1801, which can include read-only memory and random access memory, and provides operating instructions and data to the processor 1802. A portion of the memory 1801 can also include non-volatile random access memory (NVRAM).

[0320] In some embodiments, the memory 1801 stores the following elements, execution modules or data structures, or a subset thereof, or an extended set thereof.

[0321] In the embodiments of the present disclosure, by invoking the operating instructions stored in the memory 1801 (which can be stored in an operating system), the corresponding operations are performed.

[0322] Through the above description of the embodiments, those skilled in the art can clearly understand that, for the convenience and brevity of description, only the above-mentioned division of functional modules is taken as an example, and in actual application, the above-mentioned functions can be completed by different functional modules according to needs, that is, the internal structure of the device is divided into different functional modules to complete all or part of the functions described above. The specific working process of the system, device and unit described above can refer to the corresponding process in the foregoing method embodiments, which will not be described here.

[0323] Some embodiments of the present disclosure provide a computer readable storage medium (for example, a non-transitory computer readable storage medium) having computer program instructions stored therein, which, when executed on a computer (for example, a content display device), cause the computer to perform the content display method as described in any of the above embodiments.

[0324] For example, the above computer readable storage medium can include, but is not limited to, a magnetic storage device (for example, a hard disk, a floppy disk or a magnetic tape, etc.), an optical disk (for example, a CD (Compact Disk, compact disk), a DVD (Digital Versatile Disk, digital versatile disk), etc.), a smart card and a flash memory device (for example, an EPROM (Erasable Programmable Read-Only Memory, erasable programmable read-only memory), a card, a stick or a key drive, etc.). The various computer readable storage media described in the present disclosure can represent one or more devices and / or other machine readable storage media for storing information. The term "machine readable storage medium" can include, but is not limited to, a wireless channel and various other media capable of storing, containing and / or carrying instructions and / or data.

[0325] Some embodiments of the disclosure also provide a computer program product, for example, the computer program product is stored on a non-transitory computer readable storage medium. The computer program product includes computer program instructions, when the computer program instructions are executed on a computer (for example, a content display device), the computer program instructions cause the computer to perform the content display method as described in the above embodiments.

[0326] Some embodiments of the disclosure also provide a computer program. When the computer program is executed on a computer (for example, a content display device), the computer program causes the computer to perform the content display method as described in the above embodiments.

[0327] The beneficial effects of the above computer readable storage medium, computer program product and computer program are the same as those of the content display method described in some embodiments above, and will not be repeated here.

[0328] In several embodiments provided by the disclosure, it should be understood that the disclosed system, device and method can be implemented by other means. For example, the device embodiments described above are only illustrative, for example, the division of the units is only a logical function division, and actual implementation can have another division manner, for example, a plurality of units or components can be combined or integrated into another system, or some features can be ignored or not executed. In addition, the coupling or direct coupling or communication connection between the units or components shown or discussed can be indirect coupling or communication connection through some interfaces, devices or units, which can be electrical, mechanical or other forms.

[0329] The units described as separate components can or can not be physically separated, and the components shown as units can or can not be physical units, that is, they can be located in one place, or they can be distributed on a plurality of network units. Some or all of the units can be selected according to actual needs to achieve the purpose of the embodiment.

[0330] In addition, each functional unit in each embodiment of the disclosure can be integrated in one processing unit, or each unit can exist physically, or two or more units can be integrated in one unit.

[0331] The above is only a specific implementation of the disclosure, but the protection scope of the disclosure is not limited to this. Any skilled person in the art can think of changes or replacements within the technical range disclosed by the disclosure, which should be covered by the protection scope of the disclosure. Therefore, the protection scope of the disclosure should be subject to the protection scope of the claims.

Claims

1. A content display method, characterized in that, The method includes: Obtain the gaze direction of target users within the target range at the current time; the number of target users is multiple. When the target user's gaze direction at the current time is the display interface of the display device and the target user gazes at the display interface for a duration exceeding a preset duration, a user profile of the target user is determined; wherein, the user profile includes a first tag and a second tag; the first tag is used to characterize the target user's focus information, and the second tag is used to characterize the target user's feature information; For each tag in the user profile of the target user, determine the weight of each piece of content corresponding to that tag in the preset content set; For each piece of content in the preset content set, calculate the sum of the weights of each tag corresponding to that content; The content with the largest sum of weights in the preset content set is taken as the target content; Based on the content currently displayed on the display interface, control the display device to display the target content; Determining the user profile of the target user includes: Obtain the route trajectory of the target user; wherein, the route trajectory includes one or more areas passed by the target user in chronological order, and the time information of the target user's arrival at the one or more areas; The target user's target residence area is determined based on the route trajectory; the target residence area is the area in one or more regions where the target user stays for more than a preset time. The first tag is determined based on the regional information of the target residence area.

2. The method according to claim 1, characterized in that, The step of controlling the display device to display the target content based on the content currently displayed on the display interface includes: When the content currently displayed on the display interface is the first content, after the first content has been displayed, the display device is controlled to display the target content; the first content is related to the user's profile. When the content currently displayed on the display interface is the second content, the display device is controlled to stop displaying the second content and display the target content; the second content is not related to the user's profile.

3. The method according to claim 1, characterized in that, The process of obtaining the target user's route trajectory includes: Obtain the facial image information of the target user; Based on the facial image information of the target user and a first preset database, the identification information of the target user is determined; the first preset database includes multiple facial image information and identification information corresponding to each facial image information. Based on the target user's identification information, the target user's route trajectory is obtained from a second preset database; the second preset database includes the identification information of multiple users and the route trajectory of each user.

4. The method according to claim 1, characterized in that, Determining the target user's target residence area based on the route trajectory includes: Determine the target duration and target distance for the target user in the route trajectory; wherein, the target duration is the time difference between the time the target user arrives at the first area and the time the target user arrives at the second area; the target distance is the distance between the first area and the second area; the first area is any area passed by the target user, and the second area is the next area after the target user passes through the first area; If the target duration and the target distance meet the first preset conditions, the target user's target residence area is determined to be the first area.

5. The method according to claim 1, characterized in that, The feature information includes at least one of age information, gender information, facial expression information, and clothing information; When the user profile includes the second tag, determining the user profile of the target user includes: Obtain the image information of the target user; the image information of the target user includes at least one of facial image information and human body image information; The image information of the target user is input into the feature information recognition model to obtain the second label of the target user.

6. The method according to any one of claims 1-5, characterized in that, The first target tag of the target user is the tag that is the same and has the largest number of first tags among multiple target users, and the second target tag of the target user is the tag that is the same and has the largest number of second tags among multiple target users; the user profile of the target user includes at least one of the first target tag and the second target tag.

7. A content display device, characterized in that, include: The acquisition unit is configured to acquire the gaze direction of the target user within the target range at the current time; The number of target users is multiple; The processing unit is configured to determine a user profile of the target user when the target user's gaze direction at the current time is the display interface of the display device and the target user gazes at the display interface for a duration exceeding a preset duration; wherein, the user profile includes a first tag and a second tag; the first tag is used to characterize the target user's focus information, and the second tag is used to characterize the target user's feature information; The processing unit is further configured to: For each tag in the user profile of the target user, determine the weight of each piece of content corresponding to that tag in the preset content set; For each piece of content in the preset content set, calculate the sum of the weights of each tag corresponding to that content; The content with the largest sum of weights in the preset content set is taken as the target content; Based on the content currently displayed on the display interface, control the display device to display the target content; The processing unit is specifically configured as follows: Obtain the route trajectory of the target user; wherein, the route trajectory includes one or more areas passed by the target user in chronological order, and the time information of the target user's arrival at the one or more areas; The target user's target residence area is determined based on the route trajectory; the target residence area is the area in one or more regions where the target user stays for more than a preset time. The first tag is determined based on the regional information of the target residence area.

8. A content display device, characterized in that, include: Processor and memory; The memory is used to store computer programs or instructions, and the processor is used to run the computer programs or instructions to implement the content display method as described in any one of claims 1-6.

9. A computer-readable storage medium, characterized in that, The computer-readable storage medium stores instructions that, when executed by a computer, perform the content display method as described in any one of claims 1-6.

Citation Information

Patent Citations

  • Display screen, picture display method thereof and computer readable storage medium

    CN111861657A