A display method, device, apparatus and storage medium

By combining infrared sensors and cameras, the temperature and facial images of the target object are acquired, which diversifies the wake-up methods of digital humans and improves the accuracy of recognition. This solves the problems of single wake-up methods and single use of enterprise front desks in existing technologies, and saves human resources.

CN115454241BActive Publication Date: 2026-02-03SHANGHAI EGOO NETWORKS
View PDF 5 Cites 0 Cited by

Patent Information

Application Number
CN202211078515.4
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2022-09-05
Publication Date
2026-02-03
Estimated Expiration
2042-09-05

AI Technical Summary

Technical Problem

Current technologies offer limited methods for waking up digital humans, and enterprise front-end screens serve a single purpose, requiring human resources for repetitive knowledge instruction.

Method used

The infrared sensor acquires the temperature information of the target object. If the temperature is within the target temperature threshold range, the camera captures the facial image. If the image is a frontal view, the target interface is displayed. Combined with voice recognition and database comparison, diversified wake-up and accurate recognition are achieved.

Benefits of technology

It has diversified the methods of waking up digital humans, improved the accuracy of target object recognition, reduced the use of human resources, and enhanced the functional diversity of the enterprise's front end.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN115454241B_ABST
    Figure CN115454241B_ABST
Patent Text Reader

Abstract

The application discloses a display method, device, equipment and storage medium. The method comprises the following steps: obtaining temperature information of a target object through an infrared sensor; if the temperature information of the target object is within a target temperature threshold range, obtaining a facial image of the target object through a camera; and if the facial image of the target object is a front image of the target object, displaying a target interface. According to the application, the temperature information of the target object is obtained through the infrared sensor, and it is identified whether the temperature information of the target object is within the target temperature threshold range; the facial image of the target object is obtained through the camera, and it is identified whether the facial image of the target object is the front image of the target object; and then the target interface is displayed, so that the wake-up mode of the target interface is more diversified, and the identification result of the target object is more accurate.
Need to check novelty before this filing date? Find Prior Art

Description

TECHNICAL FIELD

[0001] The present application relates to the technical field of robots, and in particular to a display method, device, equipment and storage medium. BACKGROUND

[0002] In recent years, with the popularization of intelligent terminals and the development of 5G communication technology, voice recognition technology, voice synthesis technology and portrait synthesis technology have been widely applied in many industries in China. Digital people based on the above technologies are applied in intelligent customer service, live broadcast and marketing scenarios. However, the current application scenarios are limited to mobile phones or computer screens, and the wake-up method is single, mainly through the key words in the sound to wake up.

[0003] The current screen of the front desk of an enterprise has a single purpose, and is mostly used to play enterprise promotional videos or pictures. If a visiting customer has a demand, he needs to find a professional to answer and handle it. If the professional cannot explain the product or query the process, he needs to contact other staff for processing, thereby causing the problem of occupying human resources of the enterprise for repetitive knowledge explanation. SUMMARY

[0004] The present application provides a display method, device, equipment and storage medium to solve the problem of single wake-up method of digital people in the prior art, which can make the wake-up method more diversified and the recognition result of the target object more accurate.

[0005] According to an aspect of the present application, a display method is provided, which comprises:

[0006] obtaining temperature information of a target object through an infrared sensor;

[0007] if the temperature information of the target object is within a target temperature threshold range, obtaining a facial image of the target object through a camera;

[0008] if the facial image of the target object is a front image of the target object, displaying a target interface.

[0009] According to another aspect of the present application, a display device is provided, which comprises:

[0010] a first obtaining module for obtaining temperature information of a target object through an infrared sensor;

[0011] a second obtaining module for obtaining a facial image of the target object through a camera if the temperature information of the target object is within a target temperature threshold range;

[0012] a first display module for displaying a target interface if the facial image of the target object is a front image of the target object.

[0013] According to another aspect of the present application, there is provided an electronic device comprising:

[0014] at least one processor; and a memory connected with the at least one processor in communication; wherein the memory stores a computer program executable by the at least one processor, and the computer program is executed by the at least one processor to enable the at least one processor to perform the display method according to any one of the embodiments of the present application.

[0015] According to another aspect of the present application, there is provided a computer readable storage medium storing computer instructions for causing a processor to implement the display method according to any one of the embodiments of the present application when executed by the processor.

[0016] The technical solution of the embodiments of the present application acquires temperature information of a target object through an infrared sensor, acquires a facial image of the target object through a camera if the temperature information of the target object is within a target temperature threshold range, and displays a target interface if the facial image of the target object is a front image of the target object, thereby solving the problem of single digital human wake-up mode in the prior art, making the wake-up mode of the target interface more diversified, and achieving the beneficial effect of more accurate recognition result of the target object.

[0017] It should be understood that the content described in this part is not intended to identify key or important features of the embodiments of the present application, nor is it used to limit the scope of the present application. Other features of the present application will become apparent from the following description. BRIEF DESCRIPTION OF DRAWINGS

[0018] In order to more clearly illustrate the technical solutions in the embodiments of the present application, the drawings needed in the embodiment description will be briefly introduced below. Obviously, the drawings in the following description are only some embodiments of the present application, and other drawings can also be obtained by those skilled in the art without creative labor.

[0019] Figure 1 is a flow chart of a display method according to an embodiment of the present application;

[0020] Figure 2 is a structural schematic diagram of a display device according to an embodiment of the present application;

[0021] Figure 3 is a structural schematic diagram of an electronic device implementing the display method according to an embodiment of the present application. DETAILED DESCRIPTION

[0022] In the following, the technical solutions in the embodiments of the present application will be described clearly and completely with reference to the drawings in the embodiments of the present application. Obviously, the described embodiments are only a part of the embodiments of the present application, rather than all the embodiments of the present application. Based on the embodiments in the present application, all the other embodiments obtained by a person of ordinary skill in the art without creative effort should belong to the protection scope of the present application.

[0023] It should be noted that the terms "first", "target" and the like in the description, claims, and drawings of the present application are used to distinguish similar objects, and do not necessarily indicate a specific order or sequence. It should be understood that the data thus used can be interchanged under appropriate circumstances, so that the embodiments of the present application described herein can be implemented in an order other than that illustrated or described herein. In addition, the terms "include" and "have" and any variations thereof are intended to cover non-exclusive inclusion, for example, a process, method, system, product, or device that includes a series of steps or units does not necessarily have to be limited to those steps or units clearly listed, but can include other steps or units not clearly listed or inherent to the process, method, product, or device.

[0024] Embodiment One

[0025] Figure 1 is a flowchart of a display method according to Embodiment One of the present application. The present embodiment can be applied to a display case, and the method can be performed by a display device, which can be implemented in the form of hardware and / or software, and can be integrated into any electronic device that provides a display function. As shown in Figure 1 , the method comprises:

[0026] S101, obtaining temperature information of a target object by an infrared sensor.

[0027] It can be understood that the infrared sensor is a sensor that uses infrared rays for data processing, and is commonly used for non-contact temperature measurement, for example, an infrared sensor is used to remotely measure a thermal image of a human body surface temperature. In the present embodiment, the infrared sensor can be installed in an electronic screen at the front desk of an enterprise, and is used to measure the temperature of a target object within the range that the infrared sensor in front of the electronic screen can detect.

[0028] The target object can be a person, an animal, or an object, etc. within the range that the infrared sensor in front of the electronic screen can detect, and preferably, the target object can be a person within the range that the infrared sensor in front of the electronic screen can detect.

[0029] In the embodiment, the temperature information can be surface temperature information of a target object located in front of the electronic screen and within a range that can be detected by the infrared sensor.

[0030] Specifically, when the infrared sensor detects that there is a target object within the range that can be detected, the duration of the target object is started to be counted. If the duration of the target object is less than a time threshold (the time threshold can be a time set by a user according to actual conditions, and the embodiment does not limit the specific time threshold, and preferably, the time threshold can be, for example, 2 seconds), the target object is considered to be passing by. If the duration of the target object is greater than or equal to the time threshold (the time threshold can be a time set by a user according to actual conditions, and the embodiment does not limit the specific time threshold, and preferably, the time threshold can be, for example, 2 seconds), the temperature information of the target object is acquired. It should be noted that, in the embodiment, the operation of acquiring the temperature information of the target object by the infrared sensor is performed under the authorization of the user.

[0031] S102, if the temperature information of the target object is within a target temperature threshold range, a face image of the target object is acquired by the camera.

[0032] The target temperature threshold range can be a temperature range set by a user according to actual conditions, and the embodiment does not limit the specific target temperature threshold range, and preferably, the target temperature threshold range can be the normal temperature of a human body, for example, 36-37°C.

[0033] In the embodiment, the camera can be installed in the electronic screen at the front desk of the enterprise, and is used to photograph a target object within a range that can be photographed by the camera in front of the electronic screen.

[0034] It should be noted that the face image can be a face image of a target object located in front of the electronic screen and within a range that can be photographed by the camera. In actual operation, when the face image of the target object is acquired by the camera, multiple photos can be continuously taken, the acquired face image of the target object can be a front image of the target object, or a side image of the target object, and the specific number of photographs can be set by a user according to actual needs. The purpose of continuously acquiring multiple photos of the target object is to increase the recognition accuracy of the target object.

[0035] Specifically, if the temperature information of the target object obtained by the infrared sensor is within a target temperature threshold range, that is, it is determined that the target object can be a person, a camera installed in an electronic screen at the front desk of the enterprise is enabled, and a face image of the target object is obtained through the camera. It should be noted that in the embodiment, the operation of obtaining the face image of the target object through the camera is performed under the authorization of the user.

[0036] In S103, if the face image of the target object is a front image of the target object, a target interface is displayed.

[0037] It should be explained that the front image of the target object can be a face image of the target object when the target object is a person.

[0038] In the embodiment, the target interface can be an interface of a digital person displayed on an electronic screen installed at the front desk of the enterprise.

[0039] Specifically, the face image of the target object obtained by the camera is input into a preset neural network model or other face image recognition model for recognition. If it is recognized that the face image of the target object is a front image of the target object, the target interface is displayed on the electronic screen installed at the front desk of the enterprise, that is, the digital person is displayed on the electronic screen installed at the front desk of the enterprise. In actual operation, the target interface is displayed only when the obtained face image of the target object is a front image of the target object. Recognizing the front image of the target object can ensure the accuracy of the recognition of the target object, make the recognition result of the target object more accurate, and improve the recognition efficiency.

[0040] The technical scheme of the embodiment of the application obtains the temperature information of the target object through the infrared sensor. If the temperature information of the target object is within a target temperature threshold range, the face image of the target object is obtained through the camera. If the face image of the target object is a front image of the target object, the target interface is displayed. The problem of single awakening mode of the digital person in the prior art is solved. The awakening mode of the target interface can be more diversified, and the beneficial effect of more accurate recognition result of the target object is achieved.

[0041] Optionally, the temperature information of the target object is obtained through the infrared sensor, comprising:

[0042] If it is detected through the infrared sensor that there is a target object in a preset range, the staying time of the target object is obtained.

[0043] The preset range can be a range in front of the electronic screen that can be detected by the infrared sensor, or a detection range of the infrared sensor that is set by the user in advance according to the actual situation. The embodiment does not limit the specific preset range, and preferably, the preset range can be, for example, 2 meters.

[0044] It should be noted that the dwell time can be the time the target object stays within a preset range.

[0045] Specifically, if an infrared sensor detects a target object within a preset range, the time the target object stays within that preset range is obtained.

[0046] If the dwell time of the target object exceeds the time threshold, the temperature information of the target object is obtained through an infrared sensor.

[0047] The time threshold can be a time preset by the user according to the actual situation. This embodiment does not limit the specific time threshold. Preferably, the time threshold can be 2 seconds.

[0048] Specifically, if the dwell time of the target object is less than the time threshold, the target object is considered to be passing by, and the camera installed on the electronic screen at the enterprise's front desk will not be activated to obtain the target object's facial image; if the dwell time of the target object is greater than or equal to the time threshold, the target object's temperature information will be obtained, and it will be identified whether the target object's temperature information is within the target temperature threshold range.

[0049] Optionally, if the facial image of the target object is a frontal image of the target object, then the target interface is displayed, including:

[0050] The frontal image of the target object is compared with facial images in the database.

[0051] In this embodiment, the facial images in the database can be pre-registered facial images of employees within the company. It should be noted that in this embodiment, the operation of registering the facial images of employees into the database is performed with the authorization of all employees within the company.

[0052] Specifically, the frontal image of the target object captured by the camera is compared with the facial images of the company's employees pre-recorded in the database.

[0053] If the similarity value between the frontal image of the target object and the facial image in the database is greater than or equal to the similarity threshold, the first interface will be displayed.

[0054] The similarity threshold can be the similarity value between the frontal image of the target object and the facial images in the database, which is preset by the user according to the actual situation. This embodiment does not limit the specific similarity threshold. Preferably, the similarity threshold can be 80%.

[0055] In this embodiment, the first interface can be an interface displayed to employees within the enterprise. For example, the first interface may include content sections such as question answering, video playback, map search, door opening, sending messages, making phone calls, voice calls, video calls, and process queries.

[0056] Specifically, if the similarity value between the frontal image of the target object and the facial image in the database is greater than or equal to the similarity threshold, that is, if the target object is identified as possibly being an employee of the company, then the first interface, which is the interface with the access permissions of the company's internal employees, will be displayed.

[0057] Optionally, if the facial image of the target object is a frontal image of the target object, then the target interface is displayed, including:

[0058] If the similarity value between the frontal image of the target object and the facial images in the database is less than the similarity threshold, the second interface will be displayed.

[0059] The second interface is different from the first interface.

[0060] In this embodiment, the second interface can be an interface displayed to non-employees, such as external visitors. For example, the second interface may include sections for company promotion, Q&A, video playback, map search, sending messages, making phone calls, voice calls, video calls, and process queries.

[0061] Specifically, if the similarity value between the frontal image of the target object and the facial image in the database is less than the similarity threshold, that is, if the target object is identified as possibly being a non-employee of the company, such as an outside visitor, then the second interface, that is, the interface with the access rights of a stranger, is displayed.

[0062] Optionally, after displaying the first interface or the second interface, the following may also be included:

[0063] Receive voice commands input from the target object.

[0064] It should be noted that the voice commands input by the target user can be instructions given to the digital human displayed on an electronic screen installed at the company's front desk by the target user speaking. It is important to note that the digital human's operation of receiving voice commands from the target user is performed with the user's authorization. For example, the voice commands input by the target user could include opening a door, playing a company promotional video, querying the location and route to a specific place on a map, sending a message to employee A, making a phone call to employee A, or inquiring about the process of a certain matter.

[0065] Specifically, during the interaction between the target and the digital human displayed on the electronic screen installed at the enterprise's front desk, the digital human uses a voice sensor to acquire the target's speech, transcribes the target's speech into text in real time, and transmits the transcribed text to the NLP (Neuro-Linguistic Programming) component inside the electronic screen installed at the enterprise's front desk for semantic understanding, thereby acquiring the voice commands input by the target.

[0066] Determine the target operation corresponding to the voice command based on the voice command, and execute the target operation.

[0067] It should be explained that the target operation can be the operation corresponding to the voice command input by the target object. For example, if the voice command input by the target object is "open the door," then the corresponding target operation could be "open the door."

[0068] Specifically, the digital human compares the frontal image of the target object with facial images in the database. If the target object is determined to be an employee of the company, the digital human determines the target operation corresponding to the voice command based on the voice commands input by the target object, such as answering questions, playing videos, searching maps, opening doors, sending messages, making phone calls, making voice calls, making video calls, and querying processes, and then executes the target operation.

[0069] In practice, when the target interacts with the digital human displayed on an electronic screen installed at the enterprise's front desk, the digital human uses a voice sensor to capture the target's speech, transcribes the speech into text in real time, and sends the transcribed text to the NLP (Natural Language Processing) component for semantic understanding, thereby obtaining the target's voice commands. If the NLP component returns a text response, TTS (Text-to-Speech) technology is used to convert the response into audio, which is then read aloud by the digital human and displayed on the electronic screen. If the NLP component returns a video response, the digital human's video stream and the corresponding video response can be played simultaneously on the electronic screen, enabling the digital human to provide a video introduction. If the NLP component returns a map response, the location and route of the queried location on the map can be displayed on the electronic screen, while the digital human reads aloud the location and route.

[0070] In actual operation, if the user inputs the voice command "open the door," the digital human will drive the door lock to automatically open the door. The digital human is configured to be associated with the door lock; when it receives the "open the door" voice command and recognizes the frontal image of the target object as a facial image from the database, it triggers the door opening operation.

[0071] In practice, if the user's voice command is to contact employee A, the digital human can contact them via phone, video, or multimedia message. For example, when contacting by phone, the digital human can invoke the background voice call function to establish a three-way voice call between the digital human, the target, and employee A when the target tells the digital human they need to contact employee A. When contacting via voice notification, the digital human can call employee A's phone number via IVR (Interactive Voice Response) to notify employee A of the matter to be handled. When contacting via video, the digital human can initiate a video call to employee A via VoLTE (Voice over Long-Term Evolution), and after employee A answers, a three-way video call is established between the digital human, the target, and employee A. When contacting via text, the digital human can send a text message to employee A during the process of handling a matter.

[0072] In practice, digital humans can integrate with third-party OA (Office Automation) systems to perform functions such as business process query, process progress query, and process reminder. This allows for the replication of intelligent internal enterprise processes, improving workflow efficiency. It is important to note that integration with third-party OA systems is only permitted with user authorization.

[0073] In this embodiment, an image resource library can be pre-established, and videos of real people speaking and acting can be recorded and stored in the image resource library (it should be noted that the recording of videos of real people speaking and acting is performed with the user's authorization). During the training process of the digital human, after inputting text into the digital human, audio files with speaking actions can be randomly selected from the image resource library. Based on the pinyin of the input text after word segmentation, the corresponding image resource can be quickly found in the image resource library. After training is completed, the digital human can have corresponding lip movements when "speaking," and can nod, shake its head, or smile when still or "listening" to a target object, thereby improving the user experience.

[0074] Optionally, the display method also includes:

[0075] The voice input information of the target object is obtained through a sound sensor.

[0076] In this embodiment, the sound sensor can be an electronic screen installed at the front desk of the enterprise, used to acquire sound within the range that the sound sensor in front of the electronic screen can acquire.

[0077] It should be noted that voice input information can be what the target person says within the range of sound sensors on the electronic screen installed at the company's front desk.

[0078] Specifically, the sound sensor installed on the electronic screen at the company's front desk can capture voice input information from the target object within its sound range. It is important to note that this operation of capturing the target object's voice input information is performed with the user's authorization.

[0079] Recognize voice input information.

[0080] Specifically, after acquiring the voice input information of the target object through the sound sensor, the voice input information is recognized. The specific recognition process can be as follows: the digital human acquires the voice input information of the target object through the sound sensor, transcribes the content of the target object's voice input information into text in real time, and transmits the transcribed text content to the NLP component for semantic understanding, thereby completing the recognition of the voice input information.

[0081] If the voice input contains preset keywords, the target interface will be displayed.

[0082] The preset keywords can be keywords set by the user according to actual needs, which can be displayed on the electronic screen installed at the enterprise's front end. This embodiment does not limit the specific preset keywords. For example, the preset keyword can be "Hello Xiao A".

[0083] Specifically, if the voice input information contains preset keywords, the target interface will be displayed on the electronic screen installed at the enterprise's front desk, that is, a digital human will be displayed on the electronic screen installed at the enterprise's front desk; if the voice input information does not contain preset keywords, the target interface will not be displayed.

[0084] Optionally, after displaying the target interface if the voice input information contains preset keywords, the method further includes:

[0085] In response to the user's triggering of the shooting function, the camera captures a facial image of the target object.

[0086] Specifically, after recognizing preset keywords in the voice input and displaying the target interface, the system responds to the user's triggering of the camera function by acquiring a facial image of the target object through the camera. The frontal image of the target object is then compared with facial images in the database. If the similarity value between the frontal image of the target object and the facial images in the database is greater than or equal to a similarity threshold, the first interface is displayed; if the similarity value is less than the similarity threshold, the second interface is displayed. After displaying either the first or second interface, the system receives a voice command input from the target object, determines the corresponding target operation based on the voice command, and executes the target operation.

[0087] The technical solution of this invention can wake up the digital human not only through preset keywords in voice input information, but also through infrared sensors and cameras. This solves the problem of the single wake-up method of digital humans in the prior art, making the wake-up methods of the target interface more diversified, and achieving the beneficial effect of more accurate recognition results of the target object. At the same time, it solves the problem of the single use of the screen at the front desk of the enterprise and the occupation of enterprise human resources for repetitive knowledge explanation. The digital human can be displayed as an employee on the electronic screen at the front desk of the enterprise, replacing the enterprise employees in performing operations such as explaining matters, querying processes, and remote communication, which can save a lot of human resources.

[0088] Example 2

[0089] Figure 2 This is a schematic diagram of the structure of a display device according to Embodiment 2 of the present invention. Figure 2 As shown, the device includes: a first acquisition module 201, a second acquisition module 202, and a first display 203.

[0090] The first acquisition module 201 is used to acquire the temperature information of the target object through an infrared sensor.

[0091] The second acquisition module 202 is used to acquire a facial image of the target object through a camera if the temperature information of the target object is within the target temperature threshold range.

[0092] The first display module 203 is used to display the target interface if the facial image of the target object is a frontal image of the target object.

[0093] Optionally, the first acquisition module 201 includes:

[0094] The first acquisition unit is used to acquire the dwell time of the target object if the infrared sensor detects that there is a target object within a preset range.

[0095] The second acquisition unit is used to acquire the temperature information of the target object through the infrared sensor if the dwell time of the target object is greater than a time threshold.

[0096] Optionally, the first display module 203 includes:

[0097] A comparison unit is used to compare the frontal image of the target object with facial images in the database;

[0098] The first display unit is used to display a first interface if the similarity value between the frontal image of the target object and the facial image in the database is greater than or equal to a similarity threshold.

[0099] Optionally, the first display module 203 further includes:

[0100] The second display unit is used to display a second interface if the similarity value between the frontal image of the target object and the facial image in the database is less than the similarity threshold, wherein the second interface is different from the first interface.

[0101] Optionally, the first display module 203 further includes:

[0102] The receiving unit is used to receive voice commands input by the target object after the first interface or the second interface is displayed.

[0103] The processing unit is configured to, after displaying the first interface or the second interface, determine the target operation corresponding to the voice command based on the voice command, and execute the target operation.

[0104] Optionally, the display device further includes:

[0105] The third acquisition module is used to acquire the voice input information of the target object through the sound sensor;

[0106] The recognition module is used to recognize the voice input information;

[0107] The second display module is used to display the target interface if the voice input information contains preset keywords.

[0108] Optionally, the display device further includes:

[0109] The fourth acquisition module is used to acquire the facial image of the target object through a camera after displaying the target interface if the voice input information contains preset keywords, in response to the user's operation of triggering the shooting function.

[0110] The display device provided in the embodiments of the present invention can execute the display method provided in any embodiment of the present invention, and has the corresponding functional modules and beneficial effects of executing the method.

[0111] Example 3

[0112] Figure 3 A schematic diagram of an electronic device 30 that can be used to implement embodiments of the present invention is shown. The electronic device is intended to represent various forms of digital computers, such as laptop computers, desktop computers, workstations, personal digital assistants, servers, blade servers, mainframe computers, and other suitable computers. The electronic device can also represent various forms of mobile devices, such as personal digital processors, cellular phones, smartphones, wearable devices (e.g., helmets, glasses, watches, etc.), and other similar computing devices. The components shown herein, their connections and relationships, and their functions are merely illustrative and are not intended to limit the implementation of the invention described and / or claimed herein.

[0113] like Figure 3 As shown, the electronic device 30 includes at least one processor 31 and a memory, such as a read-only memory (ROM) 32 or a random access memory (RAM) 33, communicatively connected to the at least one processor 31. The memory stores computer programs executable by the at least one processor. The processor 31 can perform various appropriate actions and processes based on the computer program stored in the ROM 32 or loaded from storage unit 38 into the RAM 33. The RAM 33 can also store various programs and data required for the operation of the electronic device 30. The processor 31, ROM 32, and RAM 33 are interconnected via a bus 34. An input / output (I / O) interface 35 is also connected to the bus 34.

[0114] Multiple components in electronic device 30 are connected to I / O interface 35, including: input unit 36, such as keyboard, mouse, etc.; output unit 37, such as various types of monitors, speakers, etc.; storage unit 38, such as disk, optical disk, etc.; and communication unit 39, such as network card, modem, wireless transceiver, etc. Communication unit 39 allows electronic device 30 to exchange information / data with other devices through computer networks such as the Internet and / or various telecommunications networks.

[0115] Processor 31 can be a variety of general-purpose and / or special-purpose processing components with processing and computing capabilities. Some examples of processor 31 include, but are not limited to, a central processing unit (CPU), a graphics processing unit (GPU), various special-purpose artificial intelligence (AI) computing chips, various processors running machine learning model algorithms, a digital signal processor (DSP), and any suitable processor, controller, microcontroller, etc. Processor 31 performs the various methods and processes described above, such as display methods:

[0116] Temperature information of the target object is obtained through an infrared sensor;

[0117] If the temperature information of the target object is within the target temperature threshold range, then the facial image of the target object is acquired through the camera;

[0118] If the facial image of the target object is a frontal image of the target object, then the target interface is displayed.

[0119] In some embodiments, the display method may be implemented as a computer program tangibly contained in a computer-readable storage medium, such as storage unit 38. In some embodiments, part or all of the computer program may be loaded and / or installed on electronic device 30 via ROM 32 and / or communication unit 39. When the computer program is loaded into RAM 33 and executed by processor 31, one or more steps of the display method described above may be performed. Alternatively, in other embodiments, processor 31 may be configured to execute the display method by any other suitable means (e.g., by means of firmware).

[0120] Various embodiments of the systems and techniques described above herein can be implemented in digital electronic circuit systems, integrated circuit systems, field-programmable gate arrays (FPGAs), application-specific integrated circuits (ASICs), application-specific standard products (ASSPs), systems-on-a-chip (SoCs), payload-programmable logic devices (CPLDs), computer hardware, firmware, software, and / or combinations thereof. These various embodiments may include implementations in one or more computer programs that can be executed and / or interpreted on a programmable system including at least one programmable processor, which may be a dedicated or general-purpose programmable processor, capable of receiving data and instructions from a storage system, at least one input device, and at least one output device, and transmitting data and instructions to the storage system, the at least one input device, and the at least one output device.

[0121] Computer programs used to implement the methods of the present invention may be written in any combination of one or more programming languages. These computer programs may be provided to a processor of a general-purpose computer, a special-purpose computer, or other programmable data processing device, such that when executed by the processor, the computer programs cause the functions / operations specified in the flowcharts and / or block diagrams to be performed. The computer programs may be executed entirely on a machine, partially on a machine, or as a standalone software package, partially on a machine and partially on a remote machine, or entirely on a remote machine or server.

[0122] In the context of this invention, a computer-readable storage medium can be a tangible medium that may contain or store a computer program for use by or in conjunction with an instruction execution system, apparatus, or device. A computer-readable storage medium may include, but is not limited to, electronic, magnetic, optical, electromagnetic, infrared, or semiconductor systems, apparatus, or devices, or any suitable combination thereof. Alternatively, a computer-readable storage medium may be a machine-readable signal medium. More specific examples of machine-readable storage media include electrical connections based on one or more wires, portable computer disks, hard disks, random access memory (RAM), read-only memory (ROM), erasable programmable read-only memory (EPROM or flash memory), optical fibers, portable compact disk read-only memory (CD-ROM), optical storage devices, magnetic storage devices, or any suitable combination thereof.

[0123] To provide interaction with a user, the systems and techniques described herein can be implemented on an electronic device having: a display device (e.g., a CRT (cathode ray tube) or LCD (liquid crystal display) monitor) for displaying information to the user; and a keyboard and pointing device (e.g., a mouse or trackball) through which the user provides input to the electronic device. Other types of devices can also be used to provide interaction with the user; for example, feedback provided to the user can be any form of sensory feedback (e.g., visual feedback, auditory feedback, or tactile feedback); and input from the user can be received in any form (including sound input, voice input, or tactile input).

[0124] The systems and technologies described herein can be implemented in computing systems that include backend components (e.g., as data servers), or computing systems that include middleware components (e.g., application servers), or computing systems that include frontend components (e.g., user computers with graphical user interfaces or web browsers through which users can interact with implementations of the systems and technologies described herein), or any combination of such backend, middleware, or frontend components. The components of the system can be interconnected via digital data communication of any form or medium (e.g., communication networks). Examples of communication networks include local area networks (LANs), wide area networks (WANs), blockchain networks, and the Internet.

[0125] A computing system can include clients and servers. Clients and servers are generally located far apart and typically interact through communication networks. The client-server relationship is created by computer programs running on the respective computers and having a client-server relationship with each other. The server can be a cloud server, also known as a cloud computing server or cloud host, which is a hosting product within the cloud computing service system to address the shortcomings of traditional physical hosts and VPS services, such as high management difficulty and weak business scalability.

[0126] It should be understood that the various forms of processes shown above can be used, with steps reordered, added, or deleted. For example, the steps described in this invention can be executed in parallel, sequentially, or in different orders, as long as the desired result of the technical solution of this invention can be achieved, and this is not limited herein.

[0127] The specific embodiments described above do not constitute a limitation on the scope of protection of this invention. Those skilled in the art should understand that various modifications, combinations, sub-combinations, and substitutions can be made according to design requirements and other factors. Any modifications, equivalent substitutions, and improvements made within the spirit and principles of this invention should be included within the scope of protection of this invention.

Claims

1. A display method, characterized in that, include: Temperature information of the target object is obtained through an infrared sensor; If the temperature information of the target object is within the target temperature threshold range, then the facial image of the target object is acquired through the camera; If the facial image of the target object is a frontal image of the target object, then the target interface is displayed; If the facial image of the target object is a frontal image of the target object, then the target interface is displayed, including: The frontal image of the target object is compared with facial images in the database; If the similarity value between the frontal image of the target object and the facial image in the database is greater than or equal to the similarity threshold, the first interface is displayed; If the similarity value between the frontal image of the target object and the facial image in the database is less than the similarity threshold, a second interface is displayed. The second interface is different from the first interface. The first interface is for internal employees of the enterprise, and the second interface is for non-internal employees of the enterprise. The first interface includes sections for answering questions, playing videos, searching maps, opening doors, sending messages, making phone calls, voice calls, video calls, and process queries. The second interface includes sections for company promotion, answering questions, playing videos, searching maps, sending messages, making phone calls, voice calls, video calls, and process queries. After displaying the first interface or the second interface, it also includes: Receive voice commands input from the target object; The target operation corresponding to the voice command is determined based on the voice command, and the target operation is executed. The display method further includes: Acquire the target object's voice input information through a sound sensor; Recognize the voice input information; If the voice input information contains preset keywords, the target interface will be displayed; If the voice input information contains preset keywords, after displaying the target interface, the following steps are also included: In response to a user triggering the shooting function, the camera captures a facial image of the target object. After acquiring a facial image of the target object via the camera in response to a user triggering the shooting function, the method further includes: The frontal image of the target object is compared with facial images in the database. If the similarity value between the frontal image of the target object and the facial images in the database is greater than or equal to the similarity threshold, the first interface is displayed. If the similarity value between the frontal image of the target object and the facial images in the database is less than the similarity threshold, the second interface is displayed. After displaying the first interface or the second interface, a voice command input by the target object is received. The target operation corresponding to the voice command is determined according to the voice command, and the target operation is executed.

2. The method according to claim 1, characterized in that, Temperature information of the target object is obtained through an infrared sensor, including: If an infrared sensor detects a target object within a preset range, the dwell time of the target object is obtained. If the dwell time of the target object is greater than a time threshold, the temperature information of the target object is obtained through the infrared sensor.

3. A display device, characterized in that, include: The first acquisition module is used to acquire the temperature information of the target object through an infrared sensor; The second acquisition module is used to acquire a facial image of the target object through a camera if the temperature information of the target object is within the target temperature threshold range. The first display module is used to display the target interface if the facial image of the target object is a frontal image of the target object; The first display module includes: A comparison unit is used to compare the frontal image of the target object with facial images in the database; The first display unit is configured to display a first interface if the similarity value between the frontal image of the target object and the facial image in the database is greater than or equal to a similarity threshold. The second display unit is used to display a second interface if the similarity value between the frontal image of the target object and the facial image in the database is less than the similarity threshold. The second interface is different from the first interface. The first interface is an interface displayed to employees within the enterprise, and the second interface is an interface displayed to employees outside the enterprise. The first interface includes sections for answering questions, playing videos, searching maps, opening doors, sending messages, making phone calls, voice calls, video calls, and process queries. The second interface includes sections for company promotion, answering questions, playing videos, searching maps, sending messages, making phone calls, voice calls, video calls, and process queries. The first display module further includes: The receiving unit is used to receive voice commands input by the target object after the first interface or the second interface is displayed. The processing unit is configured to, after displaying the first interface or the second interface, determine the target operation corresponding to the voice command based on the voice command, and execute the target operation. The display device further includes: The third acquisition module is used to acquire the voice input information of the target object through the sound sensor; The recognition module is used to recognize the voice input information; The second display module is used to display the target interface if the voice input information contains preset keywords; The display device further includes: The fourth acquisition module is used to acquire the facial image of the target object through the camera after displaying the target interface if the voice input information contains preset keywords, in response to the user's operation of triggering the shooting function. After acquiring a facial image of the target object via the camera in response to a user triggering the shooting function, the method further includes: The frontal image of the target object is compared with facial images in the database. If the similarity value between the frontal image of the target object and the facial images in the database is greater than or equal to the similarity threshold, the first interface is displayed. If the similarity value between the frontal image of the target object and the facial images in the database is less than the similarity threshold, the second interface is displayed. After displaying the first interface or the second interface, a voice command input by the target object is received. The target operation corresponding to the voice command is determined according to the voice command, and the target operation is executed.

4. An electronic device, characterized in that, The electronic device includes: At least one processor; and A memory communicatively connected to the at least one processor; wherein, The memory stores a computer program that can be executed by the at least one processor, the computer program being executed by the at least one processor to enable the at least one processor to perform the display method according to any one of claims 1-2.

5. A computer-readable storage medium, characterized in that, The computer-readable storage medium stores computer instructions that, when executed by a processor, implement the display method according to any one of claims 1-2.

Citation Information

Patent Citations

  • System and method for achieving awakening and unlocking of mobile phone based on technology of temperature sense and human face recognition

    CN103167160A

  • Method and device for awakening applications

    CN104660792A

  • Intelligent mirror and control method therefor, and computer readable storage medium

    CN108776663A

  • Interface display method, device and system of operation panel and storage medium

    CN110442294A

  • Face recognition method and face recognition equipment

    CN111104818A