Prompt method, electronic device, storage medium and chip system
By establishing communication connections between online communication devices and using eye-tracking technology to obtain and prompt the other party's gaze, the problem of users not being able to know the other party's gaze is solved, thus improving the interactivity and user experience of online communication.
Patent Information
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- HUAWEI TECH CO LTD
- Filing Date
- 2024-11-29
- Publication Date
- 2026-05-29
AI Technical Summary
When users communicate with other users online, it is difficult to know what the other person is looking at, which reduces the sense of interaction.
By establishing a communication connection between the first and second devices, eye-tracking technology is used to obtain the user's gaze content, and prompts are given on the other device, such as through interface display or shape changes, to indicate that the content of the user's gaze is being paid attention to.
It enhances the interactivity of online communication, improves user experience, allows users to understand the other party's focus in real time, and enhances the authenticity and emotional resonance of communication.
Smart Images

Figure CN122111265A_ABST
Abstract
Description
Technical Field
[0001] This application relates to the field of terminal technology, and in particular to prompting methods, electronic devices, storage media, and chip systems. Background Technology
[0002] Currently, users can communicate and interact with other users through online communication methods such as video calls and video conferencing.
[0003] However, when a user communicates with other users online, it is difficult for that user to know the content corresponding to the other user's gaze point. Summary of the Invention
[0004] This application provides a prompting method, electronic device, storage medium, and chip system, which are applied in the field of terminal technology and help users understand the content corresponding to other users' gaze points when communicating with other users online.
[0005] In a first aspect, embodiments of this application provide a prompting method applied to a second device. The method includes: establishing a communication connection with a first device; and prompting the user of the first device when the user's gaze content switches to first content.
[0006] In this way, the second device can provide prompts for the content corresponding to the user's gaze point on the first device, making it easier for the user on the second device to understand the content corresponding to the user's gaze point on the first device, thereby enhancing the interactivity of online communication and improving the user experience.
[0007] In one possible implementation, the method further includes: displaying a first interface, which includes an interaction window between the first device and the second device.
[0008] The interactive window can be displayed in full screen or not; no specific restrictions are made here.
[0009] Taking full-screen interactive window display as an example, in a meeting scenario, the interactive window can be a window used for conducting the meeting, and the first interface can correspond to the following: Figure 5B , Figure 8B , Figure 12B The interface shown; in a video call scenario, the interaction window can be the video call window. The first interface can correspond to the following... Figure 14B The interface shown.
[0010] In one possible implementation, the interactive window displays first content; when the user's gaze content of the first device switches to the first content, a prompt is given for the first content, including: when the user's gaze content of the first device switches to the first content, displaying a first prompt message on the first interface, the first prompt message being used to indicate that the first content is being paid attention to; and / or, when the user's gaze content of the first device switches to the first content, the first content switches from a first form to a second form, the second form being used to indicate that the first content is being paid attention to, the first form and the second form being different.
[0011] The difference between the first and second forms can be in appearance, such as different sizes, different background colors, and different outline colors. The first prompt message can be displayed in the interactive window or outside the interactive window. There are no specific restrictions on the display location of the first prompt message.
[0012] This can be achieved by displaying prompts or by changing the format of the primary content. No specific limitations are specified here. Furthermore, it can be applied to scenarios where all content is displayed, such as a shared screen in a meeting.
[0013] In some embodiments, the first device does not display the first prompt message when the user's gaze content on the first device switches to the first content.
[0014] In one possible implementation, the interactive window displays second content, and both the second content and the first content belong to the first object; when the user's gaze content of the first device switches to the first content, a prompt is given for the first content, including: when the user's gaze content of the first device switches to the first content, a second prompt message is displayed on the first interface, the second prompt message being used to indicate that the first object is being paid attention to; and / or, when the user's gaze content of the first device switches to the first content, the second content switches from a third form to a fourth form, the fourth form being used to indicate that the first object is being paid attention to, the third form and the fourth form being different.
[0015] The differences between the third and fourth forms can be in appearance, such as different sizes, different background colors, or different outline colors. The first object can be a living or non-living entity; no specific limitation is made here. The first and second content can be the same or different. For example, the first and second content can correspond to different areas of the same object, or to the same object from different angles, or to different identifiers corresponding to the same object, etc. No specific limitation is made here.
[0016] In this way, prompts can be provided by displaying information or by changing the form of the content belonging to the first object. No specific limitations are made here. Furthermore, this approach can also be applied to scenarios where both elements display the first object, such as the same person in a virtual meeting or the same monster in a game scene.
[0017] In some embodiments, when the user's gaze content on the first device switches to the first content, the form of the first content displayed on the first device does not change.
[0018] In one possible implementation, the interactive window displays content from a first device; when the user's gaze on the first device switches to the first content, and the user's gaze on the second device switches to the third content, and the first content and the third content meet preset conditions, a prompt is given for the first content; the preset conditions include: the first content originates from the second device, and the third content originates from the first device.
[0019] This allows for notifications when users on two devices follow each other, enhancing the interactive experience. For example, a notification can be sent when a user on the first device is looking at the profile picture of a user on the second device, and vice versa. Similarly, in a meeting scenario, a notification can be sent when a user on the first device is looking at a user on the second device, and vice versa.
[0020] In one possible implementation, both the first content and the third content include features of the preset content.
[0021] The features of the preset content can be human eyes, human faces, or features of any other content. The first and third pieces of content can be the same or different.
[0022] In this way, the user can determine whether to pay attention based on the characteristics of the preset content.
[0023] In one possible implementation, the preset content includes: human eyes and / or human face.
[0024] In one possible implementation, the method further includes prompting for the fourth content when the user's gaze content on the first device switches from the first content to the fourth content.
[0025] In this way, when the gaze point changes, the content of the prompt also changes.
[0026] In one possible implementation, before prompting for the first content, the method further includes: receiving a first message when the user's gaze content of the first device switches to the first content, the first message indicating that a prompt should be made for the first content; prompting for the first content when the user's gaze content of the first device switches to the first content includes: prompting for the first content in response to the first message.
[0027] In one possible implementation, the first message is sent by the server after receiving a second message from the first device, the second message being used to indicate the first content.
[0028] In this way, the gaze content of the first device can also be obtained from the first device through the server.
[0029] In one possible implementation, the second message carries a first timestamp, which is sent by the server when the difference between the first timestamp and the time when the second message is received is less than a first preset threshold.
[0030] In this way, if the message transmission time is shorter compared to the server's current time, a notification is displayed. If the time difference is large, the message is not sent to other electronic devices, which can reduce the possibility of inaccurate notifications.
[0031] In one possible implementation, the second message carries a first timestamp, and the method further includes: the second device sending a third message to the server, the third message being used to indicate the gaze content of the user of the second device, the third message carrying a second timestamp; the first message being sent by the server when the absolute value of the difference between the first timestamp and the second timestamp is less than a second preset threshold.
[0032] In this way, if the time difference is large, the identifier of the content will not be sent to other electronic devices, which can reduce the possibility of inaccurate content prompts.
[0033] In one possible implementation, the gaze content is obtained by the first device through eye tracking.
[0034] In one possible implementation, when the user on the first device gazes at the first content for a preset duration, the user's gaze is switched to the first content. The preset duration can be 3 seconds, 2 seconds, or any other duration; no specific limitation is made here. Setting a certain duration can reduce the error in judging the user's gaze content.
[0035] Secondly, embodiments of this application provide a prompting method applied to a system, which includes a first device and a second device. The method includes: establishing a communication connection between the first device and the second device; and, when the user's gaze on the first device switches to first content, the second device provides a prompt related to the first content. In one possible implementation, the second device displays a first interface, which includes an interaction window between the first device and the second device.
[0036] In one possible implementation, the interactive window displays first content; when the user's gaze on the first content switches to the first content, the second device provides a prompt for the first content, including: when the user's gaze on the first content switches to the first content, the second device displays a first prompt message on the first interface, the first prompt message indicating that the first content is being paid attention to; and / or, when the user's gaze on the first content switches to the first content, the first content switches from a first form to a second form, the second form indicating that the first content is being paid attention to, the first form and the second form being different.
[0037] In some embodiments, the first device does not display the first prompt message when the user's gaze content on the first device switches to the first content.
[0038] In one possible implementation, the interactive window displays second content, and both the second content and the first content belong to the first object; when the user's gaze content of the first device switches to the first content, the second device provides a prompt for the first content, including: when the user's gaze content of the first device switches to the first content, the second device displays a second prompt message on the first interface, the second prompt message being used to indicate that the first object is being paid attention to; and / or, when the user's gaze content of the first device switches to the first content, the second content switches from a third form to a fourth form, the fourth form being used to indicate that the first object is being paid attention to, the third form and the fourth form being different.
[0039] In this way, prompts can be provided by displaying information or by changing the form of the content belonging to the first object. No specific limitations are made here. Furthermore, this approach can also be applied to scenarios where both elements display the first object, such as the same person in a virtual meeting or the same monster in a game scene.
[0040] In some embodiments, when the user's gaze content on the first device switches to the first content, the form of the first content displayed on the first device does not change.
[0041] In one possible implementation, the interactive window displays content from a first device; when the user's gaze on the first device switches to the first content, and the user's gaze on the second device switches to the third content, and the first content and the third content meet preset conditions, the second device provides a prompt for the first content; the preset conditions include: the first content originates from the second device, and the third content originates from the first device.
[0042] In one possible implementation, both the first content and the third content include features of the preset content.
[0043] In one possible implementation, the preset content includes: human eyes and / or human face.
[0044] In one possible implementation, the method further includes: when the user's gaze content on the first device switches from the first content to the fourth content, the second device provides a prompt for the fourth content.
[0045] In one possible implementation, before the second device prompts for the first content, the method further includes: when the user's gaze content of the first device switches to the first content, the second device receives a first message, the first message indicating that a prompt should be made for the first content; when the user's gaze content of the first device switches to the first content, the second device prompts for the first content, including: the second device responding to the first message and prompting for the first content.
[0046] In one possible implementation, the system further includes a server; the method further includes: a first device sending a second message to the server, the second message indicating the first content; and the server sending the first message to the second device after receiving the second message.
[0047] In one possible implementation, the second message carries a first timestamp. After receiving the second message, the server sends the first message to the second device, including: if the difference between the first timestamp and the time the second message was received is less than a first preset threshold, the server sends the first message to the second device.
[0048] In one possible implementation, the second message carries a first timestamp; the method further includes: the second device sending a third message to the server, the third message being used to indicate the gaze content of the user of the second device, the third message carrying a second timestamp; the server sending a first message to the second device after receiving the second message, including: the server sending the first message to the second device if the absolute value of the difference between the first timestamp and the second timestamp is less than a second preset threshold.
[0049] In this way, if the time difference is large, the identifier of the content will not be sent to other electronic devices, which can reduce the possibility of inaccurate content prompts.
[0050] In one possible implementation, before the first message, the method further includes: the first device sending an audio signal to the server, the audio signal being the user's voice collected by the first device; and the server sending the first message to the second device after receiving the second message, including: the server sending the first message to the second device after receiving the audio signal and the second message.
[0051] In this way, after the user on the first device speaks, the content corresponding to the user's gaze point on the first device is sent to the second device to prompt the user on the second device.
[0052] In one possible implementation, the method further includes: when the user's gaze content on the second device switches to the fourth content, the first device provides a prompt for the fourth content.
[0053] Thirdly, embodiments of this application provide a prompting method applied to a server. The method includes: establishing a communication connection with a first device; establishing a communication connection with a second device; and receiving a second message from the first device when the user's gaze on the first device switches to first content. The second message is used to indicate the first content; the server sends a first message to the second device after receiving the second message.
[0054] In one possible implementation, the second message carries a first timestamp. After receiving the second message, the server sends the first message to the second device, including: if the difference between the first timestamp and the time the second message was received is less than a first preset threshold, the server sends the first message to the second device.
[0055] In one possible implementation, the second message carries a first timestamp; the method further includes: the server receiving a third message from the second device, the third message being used to indicate the gaze content of the user of the second device, the third message carrying a second timestamp; the server sending a first message to the second device after receiving the second message, including: the server sending the first message to the second device if the absolute value of the difference between the first timestamp and the second timestamp is less than a second preset threshold.
[0056] In this way, if the time difference is large, the identifier of the content will not be sent to other electronic devices, which can reduce the possibility of inaccurate content prompts.
[0057] In one possible implementation, prior to the first message, the method further includes: the server receiving an audio signal from a first device, the audio signal being the user's voice collected by the first device; and the server sending the first message to the second device after receiving the second message, including: the server sending the first message to the second device after receiving the audio signal and the second message.
[0058] In this way, after the user on the first device speaks, the content corresponding to the user's gaze point on the first device is sent to the second device to prompt the user on the second device.
[0059] In one possible implementation, the method further includes: if the user's gaze content on the second device switches to the fourth content, the server sends a fifth message to the first device. The fifth message is used to indicate that a prompt should be made for the fourth content.
[0060] Thirdly, embodiments of this application provide a prompting device, which can be an electronic device, or a chip or chip system within an electronic device. The prompting device may include a display unit and a processing unit. When the interface display device is an electronic device, the display unit may be a display screen. The display unit is used to perform display steps to cause the electronic device to implement the methods described in the first aspect or any possible implementation of the first aspect. When the interface display device is an electronic device, the processing unit may be a processor. The interface display device may further include a storage unit, which may be a memory. The storage unit is used to store instructions, and the processing unit executes the instructions stored in the storage unit to cause the electronic device to implement the methods described in the first aspect or any possible implementation of the first aspect. When the interface display device is a chip or chip system within an electronic device, the processing unit may be a processor. The processing unit executes the instructions stored in the storage unit to cause the electronic device to implement the methods described in the first aspect or any possible implementation of the first aspect. The storage unit may be a storage unit within the chip (e.g., a register, cache, etc.), or a storage unit located outside the chip within the electronic device (e.g., a read-only memory, random access memory, etc.).
[0061] It is understood that the display unit is used to display an interface, such as the first interface mentioned in the first aspect above. When the user's gaze on the first device switches to the first content, the processing unit is used to provide a prompt for the first content.
[0062] Fourthly, embodiments of this application provide a system comprising: a first device and a second device. The first device is configured to perform the steps executed by the first device in the method described in the second aspect or any possible implementation of the second aspect; the second device is configured to perform the steps executed by the second device in the method described in the second aspect or any possible implementation of the second aspect.
[0063] Fifthly, embodiments of this application provide an electronic device including a processor and a memory, the memory for storing code instructions, and the processor for executing the code instructions to perform the methods described in the first aspect or any possible implementation of the first aspect.
[0064] In a sixth aspect, embodiments of this application provide a computer-readable storage medium storing a computer program or instructions that, when executed on a computer, cause the computer to perform the methods described in the first aspect or any possible implementation thereof.
[0065] In a seventh aspect, embodiments of this application provide a computer program product including a computer program, which, when run on a computer, causes the computer to perform the methods described in the first aspect or any possible implementation thereof.
[0066] Eighthly, this application provides a chip or chip system including at least one processor and a communication interface, the communication interface and the at least one processor being interconnected via a circuit, the at least one processor being used to run computer programs or instructions to perform the methods described in the first aspect or any possible implementation thereof. The communication interface in the chip may be an input / output interface, pins, or circuits, etc.
[0067] In one possible implementation, the chip or chip system described above in this application further includes at least one memory storing instructions. The memory can be an internal storage unit of the chip, such as a register or cache, or it can be a storage unit of the chip itself (e.g., read-only memory, random access memory, etc.).
[0068] It should be understood that the second to eighth aspects of this application correspond to the technical solutions of the first aspect of this application, and the beneficial effects achieved by each aspect and the corresponding feasible implementation are similar, and will not be repeated here. Attached Figure Description
[0069] Figure 1 This is a flowchart illustrating one possible vision correction method in a design.
[0070] Figure 2A flowchart illustrating one embodiment of the prompting method provided in this application;
[0071] Figure 3A A flowchart illustrating an eye-tracking method provided in an embodiment of this application;
[0072] Figure 3B This is a schematic diagram showing the correspondence between controls and control hotspots provided in the embodiments of this application;
[0073] Figure 3C A schematic diagram of an interface for control recognition provided in an embodiment of this application;
[0074] Figure 3D A schematic diagram of a text recognition interface provided for an embodiment of this application;
[0075] Figure 3E A schematic diagram of a graphic recognition interface provided in an embodiment of this application;
[0076] Figure 4A This is a schematic diagram illustrating a scenario where a prompt is displayed when multiple electronic devices are displaying the same screen, as provided in an embodiment of this application.
[0077] Figure 4B A schematic diagram of the interface of an electronic device B provided in an embodiment of this application;
[0078] Figure 4C A schematic diagram of the interface of an electronic device C provided in an embodiment of this application;
[0079] Figure 5A This is another scenario illustration provided for an embodiment of this application;
[0080] Figure 5B A schematic diagram of the interface of various electronic devices provided in the embodiments of this application;
[0081] Figure 6 A flowchart illustrating another embodiment of the prompting method provided in this application;
[0082] Figure 7A A schematic diagram of a video conferencing scenario provided for an embodiment of this application;
[0083] Figure 7B A schematic diagram of the interface of various electronic devices provided in the embodiments of this application;
[0084] Figure 8A A schematic diagram of a video conferencing scenario provided for an embodiment of this application;
[0085] Figure 8B A schematic diagram of the interface of various electronic devices provided in the embodiments of this application;
[0086] Figure 9 A flowchart illustrating a content overlap determination method provided in an embodiment of this application;
[0087] Figure 10A A schematic diagram illustrating a scenario where control hotspots overlap and do not overlap, provided for an embodiment of this application;
[0088] Figure 10B This is a schematic diagram illustrating a scenario where text overlaps and does not overlap, provided for an embodiment of this application.
[0089] Figure 10C This is an illustration of another scenario where text overlaps, provided as an embodiment of this application.
[0090] Figure 10D A schematic diagram illustrating a scenario of overlapping and non-overlapping graphics provided in an embodiment of this application;
[0091] Figure 11 A flowchart illustrating a time determination method provided in an embodiment of this application;
[0092] Figure 12A A schematic diagram of a virtual meeting scenario provided for an embodiment of this application;
[0093] Figure 12B A schematic diagram of the interfaces of various electronic devices in a virtual meeting scenario provided in an embodiment of this application;
[0094] Figure 12C A schematic diagram of the interfaces of various electronic devices in a virtual meeting scenario provided in an embodiment of this application;
[0095] Figure 13 A flowchart illustrating a prompting method provided in an embodiment of this application;
[0096] Figure 14A This is a schematic diagram of an application scenario provided by an embodiment of this application;
[0097] Figure 14B This is a schematic diagram of the interfaces of various electronic devices in a video call scenario provided in an embodiment of this application;
[0098] Figure 15 A flowchart illustrating a prompting method provided in an embodiment of this application;
[0099] Figure 16 This is a schematic diagram of the structure of a prompting device provided in an embodiment of this application;
[0100] Figure 17 This is a schematic diagram of the structure of an electronic device provided in an embodiment of this application. Detailed Implementation
[0101] For the convenience of understanding, the relevant terms and concepts involved in the embodiments of the present application are introduced below:
[0102] 1. Eye movement heat map
[0103] The eye movement heat map is a visualization tool used to display the fixation behavior of a user when viewing a certain visual scene (such as a web page, advertisement, image, or video). In the embodiments of the present application, the electronic device can record and analyze the eye movement data to obtain the areas of interest and the intensity of fixation of the user, and display them through the eye movement heat map.
[0104] The heat map usually uses color coding to represent the frequency and duration of fixation. For example, red indicates a high-interest area (long-time fixation or multiple fixations), and blue or green indicates a less-concerned area.
[0105] 2. Electronic device
[0106] The electronic device in the embodiments of the present application can be referred to as a user equipment (UE), a terminal, etc. For example, the electronic device can be a mobile phone, a tablet computer, a personal digital assistant (PDA), a handheld device with wireless communication function, a computing device, a vehicle-mounted device, or a wearable device, a virtual reality (VR) terminal device, an augmented reality (AR) terminal device, a wireless terminal in industrial control, a wireless terminal in a smart home, etc. In the embodiments of the present application, the form of the electronic device is not specifically limited.
[0107] 3. Other terms
[0108] In the embodiments of the present application, "at least one" means one or more, and "multiple" means two or more. "And / or" describes the association relationship of associated objects and indicates that three relationships can exist. For example, A and / or B can represent: A exists alone, A and B exist simultaneously, and B exists alone, where A and B can be singular or plural. The character " / " generally indicates that the associated objects before and after are in an "or" relationship. "At least one (item)" or its similar expression refers to any combination of these items, including any combination of single item (item) or plural items (items). For example, at least one (item) of a, b, or c can represent: a, b, c, a - b, a - c, b - c, or a - b - c, where a, b, and c can be single or multiple.
[0109] In interpersonal interactions, eye contact is real-time and very important. However, when users communicate with other users through video calls, virtual spaces, or remote video, it is difficult for users to know whether they have made real-time eye contact with other users or whether their focus is on the same content, so as to obtain resonance of emotions, intentions, and / or attitudes.
[0110] In a possible design, during a video chat, the electronic device can correct the user's gaze in the image of the human eye, so that the gaze in the image of the human eye is direct.
[0111] Specifically, such as Figure 1 As shown, the electronic device can acquire a color image through a color camera and a depth image aligned with the color image through a depth camera. The electronic device performs preprocessing on the color and depth images to obtain 3D information in a virtual coordinate system, and projects this information onto the virtual camera plane to obtain an image with corrected eye contact. The electronic device then applies median filtering to the corrected eye contact image to obtain an image showing direct eye contact. In this way, by correcting the user's gaze in the video image through an algorithm, it is possible to perceive direct eye contact even when the user is not looking at the camera.
[0112] However, even when the user looks elsewhere, the eyes of the people in the video remain fixed on the same point, failing to represent the user's actual, real-time eye contact.
[0113] In some embodiments, electronic devices can use eye-tracking devices to enable user interaction with the interface or predict and analyze user attention. Specifically, electronic devices can use eye-tracking devices (e.g., IR or RGB cameras) and algorithms such as pupil-corneal algorithms and fixation algorithms to track the user's gaze in real time. By analyzing the area corresponding to the user's gaze and the intensity of the gaze, interface control operations or user attention prediction and analysis can be achieved. For example, clicking a button when the user gazes at an interface button, or generating an eye-tracking heatmap to determine the user's focus point.
[0114] In view of this, embodiments of this application provide a prompting method, an electronic device, a storage medium, and a chip system. Each electronic device can obtain the content corresponding to the gaze point of the corresponding user through its corresponding eye-tracking device and send the identifier of the content corresponding to the gaze point of its corresponding user to the server. The server can send the identifiers of the content corresponding to the gaze points of users on other devices to the electronic devices. Subsequently, the electronic devices can prompt the user based on the identifiers to facilitate the user's understanding of the content corresponding to the gaze points of users on other devices, thereby improving the interactivity of online communication and enhancing the user experience.
[0115] The methods and apparatus provided in this application can be applied to a communication system, which may include a server, a first device, and a second device. Both the first device and the second device establish a communication connection with the server.
[0116] In this embodiment, the first device performs eye tracking to obtain the content corresponding to the user's gaze point on the first device. The server can obtain the identifier of the content corresponding to the user's gaze point on the first device and transmit the identifier to the second device. The second device can provide prompts to the user based on the content identifier. The second device can provide prompts to the user in one or more ways, such as displaying a prompt box on the interface, displaying animation effects, playing voice, vibration, etc., without specific limitations here.
[0117] In some embodiments, the second device can also be used to perform eye tracking to obtain the content corresponding to the user's gaze point of the second device; the server can also obtain the content identifier corresponding to the user's gaze point of the second device and transmit the content identifier corresponding to the user's gaze point of the second device to the first device; the first device can also provide prompts to the user of the first device based on the content identifier.
[0118] In this embodiment, the server can be an independent physical server, a server cluster or distributed system composed of multiple physical servers, or a cloud server that provides basic cloud computing services such as cloud services, cloud databases, cloud computing, cloud functions, cloud storage, network services, cloud communication, middleware services, domain name services, security services, content delivery networks (CDN), and big data and artificial intelligence platforms.
[0119] Both the first device and the second device can be any type of electronic device, as detailed in the above descriptions, which will not be elaborated upon here. The types of the first device and the second device can be the same or different, and no specific restrictions are imposed here.
[0120] Both the first and second devices establish a communication connection with the server. This communication connection can be of any form, such as any form of wireless communication or any form of wired communication. For example, a wireless communication connection can include the following forms: Wireless Fidelity (WIFI) communication connection, cellular communication connection, Bluetooth communication connection, near-field communication connection, ultra-wide band (UWB) communication connection, and infrared transmission communication connection, etc. A wired communication connection can include the following forms: Ethernet communication connection, fiber optic communication connection, and universal serial bus (USB) connection, etc.
[0121] In this embodiment of the application, communication connections between various electronic devices and servers can be established through protocols such as Hypertext Transfer Protocol (HTTP) or Hypertext Transfer Protocol over Secure Socket Layer (HTTPS). This embodiment of the application does not impose any restrictions on this.
[0122] The above embodiments are illustrated using the example of a communication system comprising two devices and a server. The communication system may also include more or fewer electronic devices.
[0123] The technical solution of this application and how it solves the above-mentioned technical problems are described in detail below with specific embodiments. These specific embodiments can be implemented independently or in combination with each other. The same or similar concepts or processes may not be described again in some embodiments.
[0124] The following is combined with Figures 2 to 14B This section explains the application of prompting methods in different scenarios. Figures 2 to 11 This can correspond to the application of prompting methods in scenarios where the same screen is displayed on various electronic devices; Figures 12A to 14B This can correspond to the application of prompting methods in scenarios where different screens are displayed on various electronic devices.
[0125] For example, Figure 2 This is a flowchart illustrating a notification method provided in an embodiment of this application. Taking, for example, devices A and B that access the system via a server. Figure 2 As shown, the method includes:
[0126] S201. After connecting to the video conference, electronic device A performs eye tracking to obtain the content A corresponding to user A's gaze point.
[0127] User A corresponds to electronic device A, and can also be understood as the user using electronic device A. In some embodiments, after detecting that user A is gazing at content A for a preset duration, the content A corresponding to user A's gaze point is obtained. The preset duration can be 3 seconds, 2 seconds, or any duration, and is not specifically limited here.
[0128] In this embodiment, electronic device A is equipped with an eye-tracking device. The electronic device can perform eye tracking through the eye-tracking device.
[0129] Eye-tracking devices include, but are not limited to: eye trackers, virtual reality wearable devices, infrared (IR) cameras, or red-green-blue (RGB) cameras.
[0130] For example, such as Figure 3A As shown, S201 may include: S2011-S2013.
[0131] S2011, The eye-tracking device acquires the physical gaze point of user A.
[0132] The physical fixation point can be understood as the actual location where the eye is focused in physical space. This location is usually a point defined in three-dimensional space.
[0133] In this embodiment, the eye-tracking device can acquire the physical gaze point of user A in any manner. For example, the eye-tracking device can acquire multiple face images, identify and track pupil and corneal reflections (typically "bright spots" generated by an infrared light source) from the face images. Based on the position and relative movement of these features, the gaze direction is calculated to obtain the physical gaze point of user A. In some embodiments, the eye-tracking device can periodically acquire the physical gaze point of user A (e.g., periodically acquire and analyze face images). As an example, and not a limitation, for instance, if the frequency of image acquisition by the eye-tracking device is 30 Hz, that is, approximately every 33 milliseconds (ms), then the acquisition period is 33 ms, and the period for acquiring the physical gaze point of user A is 33 ms. S2012, the electronic device A converts the physical gaze point into an interface gaze point.
[0134] The interface gaze point can be understood as the specific location where a user is looking at the screen of an electronic device or other display device. This point is usually defined on a two-dimensional plane and represents the user's visual focus on the interface.
[0135] Converting a physical gaze point to an interface gaze point can be understood as mapping the gaze point in physical space to the coordinate system of the electronic device's display device (e.g., a screen) or virtual screen. For example, the process of converting a physical gaze point to an interface gaze point by electronic device A may include coordinate transformation, viewpoint correction, data filtering, etc. No specific limitations are made here.
[0136] Coordinate transformation maps the gaze point in physical space to the coordinate system of an electronic device's display (e.g., a screen) or virtual screen, facilitating the subsequent determination of the content on the interface corresponding to the gaze point. Viewpoint correction reduces the impact of user head and eye movements on the gaze point's position, improving the accuracy of gaze point localization on the electronic device's display (e.g., a screen) or virtual screen. Data filtering reduces noise interference in eye-tracking data, further improving the accuracy of gaze point localization on the electronic device's display (e.g., a screen).
[0137] S2013, Electronic device A recognizes the content corresponding to the gaze point on the interface.
[0138] Electronic device A can obtain the content corresponding to the interface gaze point by comparing the position of each element (e.g., controls, text, images, etc.) in the interface displayed by the electronic device.
[0139] Controls can be understood as elements in a human-computer interface (HCI) that users can interact with through touch, click, etc., such as buttons, text links, cards, and icon buttons. Text can be understood as readable text in a HCI. Graphics can be understood as visual graphics visible to the user in a HCI.
[0140] In some embodiments, if electronic device A recognizes the content corresponding to the interface gaze point as content A in N consecutive instances, the content A corresponding to user A's gaze point is obtained.
[0141] It should be noted that the text displayed on the interface can exist in text form, or in the form of images, controls, etc. In this embodiment, text recognition refers to the recognition of text in text form. Text in image form can be recognized as a graphic using the following graphic recognition method. Text in control form can be recognized as a control using the following control recognition method. The recognition of each element can be referred to the following... Figures 9 to 10D The corresponding description is not detailed here.
[0142] Specifically, when the gaze point area and the corresponding control hotspot of any control meet preset condition A, the identified content includes: the control corresponding to that control hotspot. When the gaze point area and any text in the interface meet preset condition B, the identified content includes: the text; when the gaze point area and any graphic meet preset condition C, the identified content includes: the graphic.
[0143] A control hotspot can be understood as an invisible area of a specific size attached to a control, within which the user needs to interact to trigger a control response. The size of the control hotspot can be larger or smaller than the size of the control; the shape of the control hotspot can be the same as or different from the shape of the control. This application does not impose specific limitations on the shape, size, etc., of the control hotspot.
[0144] For example, Figure 3B This is a schematic diagram showing the corresponding control and control hotspot provided in the embodiments of this application. Figure 3B Interface 3B, shown in figure 'a', includes multiple controls, such as control 301, control 302, and control 303; the control hotspots corresponding to each control can be configured as follows: Figure 3B As shown in b in the figure. For example, the control hotspot corresponding to control 301 is gray area 304, the control hotspot corresponding to control 302 is gray area 305, and the control hotspot corresponding to control 303 is gray area 306.
[0145] In some embodiments, preset condition A may include an area of overlap between the gaze point region and the control hotspot that is greater than or equal to a threshold A. Preset condition B may include an area of overlap between the gaze point region and text that is greater than or equal to a threshold B. Preset condition C may include an area of overlap between the gaze point region and graphics that is greater than or equal to a threshold C.
[0146] Threshold A can be the fixation point region, half of the fixation point region, or any value; threshold B can be half of a single character, a single character, half of a word, or any value; threshold C can be the fixation point region, half of the fixation point region, half of a graphic, or any value, etc. This application does not limit the specific values of threshold A, threshold B, and threshold C.
[0147] It should be noted that the values of thresholds A, B, and C can be related to the specific application scenario. For example, taking threshold C as an example, in scenarios with high precision requirements (e.g., recognizing eyes in an image during a two-way video call), threshold C is larger, for example, 50% of the image area; in scenarios with lower precision requirements (e.g., recognizing eyes in an image during a meeting), threshold C is smaller, for example, 10% or 1% of the image area. No specific limitations are imposed here.
[0148] For example, taking a threshold A as half of the fixation point region as an example, such as... Figure 3C As shown, if the gaze point area in interface 3B is the black area 307, the content obtained does not include: control 301 corresponding to the gray area 304; if the gaze point area in interface 3B is the black area 308 or the gray area 309, the content obtained includes: control 302 corresponding to the gray area 304.
[0149] For example, taking threshold B as a single character, such as... Figure 3D As shown in 'a', if the fixation point region is gray area 310, the obtained content includes: "Threshold A Threshold B". If the fixation point region is gray area 311, the obtained content includes: "Dot area text".
[0150] If the threshold B is half of the fixation point region, such as Figure 3D As shown in b, if the fixation point region is gray area 312, the obtained content includes: "A is" corresponding to the entire text "Threshold A is half of the fixation point region". If the fixation point region is gray area 313, the obtained content includes: "text" corresponding to the entire text "Threshold B is a single text".
[0151] For example, taking a threshold C as half the size of the graph as an example, such as... Figure 3E As shown, if the fixation point region is gray area 314, the obtained content includes: an eye graphic. If the fixation point region is gray area 315, the obtained content does not include: an eye graphic.
[0152] In addition to the area of the overlapping region mentioned above, the content corresponding to the gaze point can also be identified in any other way, such as when there is an overlap between the gaze point region and the control hotspot, when the edge of the gaze point region is located within the control hotspot, or when the edge of the gaze point region is located within the graphic. No specific limitations are made here.
[0153] In this embodiment, electronic device A may begin eye tracking after connecting to a system (e.g., a video conferencing system, a video call system); it may also begin eye tracking after receiving a control indicating to turn on the camera; or it may begin eye tracking after electronic device A acquires an audio signal. This embodiment does not specifically limit the triggering method for eye tracking.
[0154] S202. Electronic device A sends message A to the server. Message A includes: identifier A corresponding to content A.
[0155] The identifier A is used to identify content A. Taking content A as an image as an example, the identifier A can be the name of the image, the name of the component corresponding to the image, or the image itself.
[0156] S203. After receiving message A, the server sends message B to electronic device B. Message B includes: identifier A.
[0157] In some embodiments, message B further includes an identifier indicating the speaker's gaze. This allows other electronic devices to subsequently use the information in message A to indicate the content corresponding to the speaker's gaze.
[0158] S204. After receiving message B, electronic device B displays the content A corresponding to identifier A.
[0159] In this embodiment of the application, the prompting method can be voice prompt, display prompt, vibration prompt, etc., and no specific limitation is made here.
[0160] It should be understood that the embodiments in this application are illustrated using the identification of content transmitted between devices as an example, and content can also be transmitted between devices. No specific limitations are made here.
[0161] The following is combined with Figures 4A to 14B This section explains the application of prompting methods in different scenarios. Figures 4A to 9 This can correspond to the application of prompting methods in scenarios where the same screen is displayed on various electronic devices; Figures 11 to 14B This can correspond to the application of prompting methods in scenarios where different screens are displayed on various electronic devices.
[0162] The following is combined with Figure 4A The video conference scene shown, and Figure 4B and Figure 4C The interface shown Figure 2 The prompts shown are explained below.
[0163] For example, Figure 4A This is a schematic diagram illustrating a scenario where multiple electronic devices display the same screen, as provided in an embodiment of this application, and a prompt is displayed. Figure 4A The scenario shown includes: electronic device A, electronic device B, user A, user B, and server 401. User A is within the detection range of the eye-tracking device on electronic device A. Electronic devices A and B display the same screen.
[0164] When user A gazes at the display screen of electronic device A, electronic device A can detect user A's gaze point on the display screen through an eye-tracking device, and then obtain the content corresponding to user A's gaze point.
[0165] Taking the identifier A corresponding to the content at user A's gaze point as an example, electronic device A can transmit identifier A to electronic device B through server 401. After receiving identifier A, electronic device B can annotate content A on its interface.
[0166] like Figure 4B As shown, electronic device B can display a prompt message 402 on the interface. The prompt message 402 is used to indicate the area corresponding to the gaze point of electronic device B. For example, the prompt message 402 could be "User A is looking at this area".
[0167] In some embodiments, the prompt information 402 is located near the content corresponding to identifier A and points to that content. This makes it easier for the user to confirm the content corresponding to the user's gaze point on electronic device A.
[0168] The above Figures 4A to 4C The added prompt information shown is only one way to display prompts. Electronic devices can also provide prompts in any other way, such as changing the background color of the content corresponding to identifier A, modifying the size of the content corresponding to identifier A (e.g., enlarging it), modifying the color of the content corresponding to identifier A, displaying animation effects, etc. There are no specific limitations here.
[0169] In summary, electronic device B can obtain the content corresponding to the user's gaze point on electronic device A, and prompt the user of electronic device B in one or more ways, so that the user can understand the content corresponding to the user's gaze point on electronic device A, thereby improving the interactivity, communication efficiency, accuracy, immersion and fun of video conferencing, and enhancing the user experience.
[0170] Building upon the above embodiments, if the devices accessing the video conference via the server include additional devices, the server can also send message B to other devices in the video conference. In this way, each participant's electronic device can receive prompts based on the content corresponding to the speaker's gaze, allowing the participant to understand the content corresponding to the user's gaze on electronic device A. This enhances the interactivity, communication efficiency, accuracy, immersion, and engagement of the video conference, thereby improving the user experience.
[0171] For example, the device for accessing a video conference via a server also includes electronic device C. The interface method further includes: the server sending message B to electronic device C. Message B includes: an identifier A. After receiving message B, electronic device C displays a prompt for content A corresponding to identifier A.
[0172] For example, with Figure 4A The video conferencing scenario shown also includes electronic device C as an example. The interface displayed on electronic device C can be as follows: Figure 4C As shown. Specifically, electronic device C can display a prompt message 402 on the interface. The prompt message 402 is used to indicate the area corresponding to the gaze point of electronic device C. For example, the prompt message 402 could be "User A is looking at this area".
[0173] It is understandable that the way electronic device C prompts content A can be the same as or different from the way electronic device B prompts it; no specific limitation is made here.
[0174] The above embodiment uses electronic device A for eye tracking and other devices to provide prompts for the content corresponding to user A's gaze point. In a video conferencing scenario, the electronic device can also provide prompts for the content corresponding to the gaze points of users on other electronic devices.
[0175] For example, Figure 5A This is a schematic diagram of another video conferencing scenario provided for an embodiment of this application. Figure 5A The scenario shown includes: electronic device A, electronic device B, electronic device C, user A, user B, user C, and server 501. User A is within the detection range of the eye-tracking device on electronic device A, user B is within the detection range of the eye-tracking device on electronic device B, and user C is within the detection range of the eye-tracking device on electronic device C. Electronic devices A, B, and C display the same screen.
[0176] Taking the identifier of the content corresponding to user A's gaze point as identifier A, the identifier of the content corresponding to user A's gaze point as identifier B, and the identifier of the content corresponding to user A's gaze point as identifier C as an example, electronic device A can transmit identifier A to electronic devices B and C through server 501; electronic device B can transmit identifier B to electronic devices A and C through server 501; and electronic device C can transmit identifier C to electronic devices A and B through server 501.
[0177] Adaptably, electronic device A can annotate content B corresponding to identifier B and content C corresponding to identifier C on the interface. Electronic device B can annotate content B corresponding to identifier B and content C corresponding to identifier C on the interface. Electronic device A can annotate content B corresponding to identifier B and content C corresponding to identifier C on the interface.
[0178] like Figure 5B As shown, electronic device A displays prompts 502 and 503 on its interface. Electronic device B displays prompts 503 and 504 on its interface. Electronic device C displays prompts 502 and 504 on its interface. Prompt 502 indicates the area corresponding to the user's gaze point on electronic device B. For example, prompt 502 could be "User B is paying attention here." Prompt 503 indicates the area corresponding to the user's gaze point on electronic device C. For example, prompt 503 could be "User C is paying attention here." Prompt 504 indicates the area corresponding to the user's gaze point on electronic device A. For example, prompt 504 could be "User A is paying attention here." This embodiment of the application does not specifically limit the content, location, etc., of each prompt.
[0179] In some embodiments, the server can statistically analyze the content corresponding to the user's gaze points on each electronic device to obtain statistical results. The electronic devices can then provide prompts based on these statistical results.
[0180] In one possible implementation, the server could statistically analyze the content corresponding to user gaze points on each electronic device, obtaining the number of users corresponding to each identifier. Subsequently, when each electronic device displays prompts for the content of each identifier, it could also indicate the number of people gazing at that content, the probability value, etc. The probability value could be the ratio between the number of users and the total number of users. In this way, users can understand the number of people gazing at each part of the content, the probability of that gaze, etc.
[0181] In some embodiments, the electronic device can also determine, based on statistical results, whether each electronic device should provide prompts for content corresponding to the gaze points of other devices. This reduces the need for prompts and makes it easier for users to understand the content corresponding to the gaze points of most people.
[0182] For example, an electronic device can display a prompt for the content corresponding to an identifier when the number of users corresponding to that identifier is greater than or equal to a threshold A. The electronic device can also display a prompt for the content corresponding to an identifier when the number of users corresponding to that identifier is less than the threshold A.
[0183] The electronic device can provide a prompt when the probability value corresponding to the identifier is greater than or equal to a threshold B. The electronic device can also provide a prompt when the probability value corresponding to the identifier is less than the threshold B.
[0184] This application does not limit the specific implementation of electronic devices providing prompts based on statistical results, nor does it limit the prompting method.
[0185] Based on the above embodiments, the electronic device can also determine whether to provide prompts for the content corresponding to the user's gaze point of the electronic device according to the category of each electronic device.
[0186] The embodiments of this application can classify electronic devices based on one or more methods, such as whether a user's voice signal is detected (which can also be understood as whether the user corresponding to the electronic device speaks), user-defined settings, etc.
[0187] Taking user-defined settings as an example, users can designate all or some of the other electronic devices in the communication system as the devices that require prompts. This allows users to freely configure the prompting for each device based on the user's gaze point, personalizing the settings and improving the user experience. Furthermore, it can reduce the number of prompts in video conferences, minimizing distractions.
[0188] For example, after a user sets up prompts for content corresponding to the user's gaze points on certain devices, the server can send an identifier to the user's corresponding electronic device indicating the content to be transmitted by the electronic device that is set to provide prompts, but not send an identifier to the user's corresponding electronic device indicating that no electronic device is set to provide prompts for such content. This reduces the signaling interaction between the server and various electronic devices.
[0189] For example, after a user sets up prompts for content corresponding to the user's gaze point on certain devices, the server can send the identifiers of the content transmitted by each electronic device and the corresponding user identifiers to the user's corresponding electronic device. The user's corresponding electronic device can prompt the content corresponding to that user identifier if the user identifier corresponds to an electronic device set up for prompting; otherwise, it will not prompt the content corresponding to that user identifier.
[0190] Taking the detection of a user's voice signal as an example, the electronic devices in a communication system can be divided into: the electronic device corresponding to the speaker and the electronic device corresponding to the participant.
[0191] For example, the server can send an identifier of the content corresponding to the speaker's gaze point to each participant's electronic device, so that each participant's device can provide prompts for the content corresponding to the speaker's gaze point. This makes it easier for each participant to understand the content corresponding to the speaker's gaze point. The server can also send an identifier of the content corresponding to each participant's gaze point to each speaker's electronic device, so that each speaker's electronic device can provide prompts for the content corresponding to each participant's gaze point. This makes it easier for the speaker to understand the content corresponding to each participant's gaze point.
[0192] For example, Figure 6 This is a flowchart illustrating a notification method provided in an embodiment of this application. Taking, for example, devices accessing a video conference via a server, including electronic device A, electronic device B, and electronic device C, as shown... Figure 6 As shown, the method includes:
[0193] S601. Electronic device A sends user A's voice signal to the server.
[0194] In this embodiment of the application, electronic device A can collect the user's voice through a microphone or other device.
[0195] S602. After receiving the voice signal from user A, the server sets electronic device A as the speaker's electronic device.
[0196] S603. After connecting to the video conference, electronic device A performs eye tracking to obtain the content A corresponding to user A's gaze point.
[0197] S604. Electronic device A sends message A to the server. Message A includes: identifier A corresponding to content A.
[0198] S605. After connecting to the video conference, electronic device B performs eye tracking to obtain the content B corresponding to user B's gaze point.
[0199] S606. Electronic device B sends message B to the server. Message B includes: identifier B corresponding to content B.
[0200] S607. After connecting to the video conference, electronic device C performs eye tracking to obtain the content C corresponding to the user C's gaze point.
[0201] S608. Electronic device C sends message C to server. Message C includes: identifier C corresponding to content C.
[0202] The order of S603, S605, and S607 is not limited in the embodiments of this application.
[0203] S609. The server sends message D to electronic device A. Message D includes: identifier B and identifier C.
[0204] In some embodiments, the server may also send identifier B and identifier C separately, which is not specifically limited here.
[0205] S610, Electronic device A provides prompts for content B and content C.
[0206] S611. The server sends message E to electronic device B. Message E includes: identifier A.
[0207] S612. Electronic device B provides a prompt for content A.
[0208] S613. The server sends message F to electronic device C. Message F includes: identifier A.
[0209] S614. Electronic device C provides a prompt for content A.
[0210] In this way, each participant's electronic device can provide prompts for the content corresponding to the speaker's gaze point; the speaker's electronic device can provide prompts for the content corresponding to each participant's gaze point.
[0211] For example, Figure 7A This is a schematic diagram of a video conferencing scenario and the interfaces of various electronic devices provided in an embodiment of this application. Figure 7AThe scenario shown includes: electronic device A, electronic device B, electronic device C, user A, user B, user C, and server 701. Users A, B, and C are all within the detection range of their respective electronic devices' eye-tracking devices. Electronic devices A, B, and C display the same screen.
[0212] Taking user A as the speaker as an example, such as Figure 7B As shown, electronic device A displays prompt information 702 and prompt information 703 on the interface. Electronic device B displays prompt information 704 on the interface. Electronic device C displays prompt information 704 on the interface. Prompt information 702 is used to indicate the area corresponding to the user's gaze point on electronic device B. For example, prompt information 702 could be "User B is paying attention here." Prompt information 703 is used to indicate the area corresponding to the user's gaze point on electronic device C. For example, prompt information 703 could be "User C is paying attention here." Prompt information 704 is used to indicate the area corresponding to the speaker's gaze point. For example, prompt information 704 could be "The speaker is paying attention here." This application embodiment does not specifically limit the content, location, etc., of each prompt information.
[0213] In the above Figure 6 Based on the illustrated embodiment, the server can also send identifiers of the content corresponding to the gaze points of other participants to the electronic devices of each participant, so that each participant's device can provide prompts for the content corresponding to the gaze points of other participants. The specific implementation is the same as described above. Figure 2 The process shown is similar, and will not be described in detail here.
[0214] In some embodiments, a prompt is given when a large number of attendees are looking at the same content; no prompt is given when a small number of attendees are looking at the same content.
[0215] Based on the above embodiments, the server can also confirm whether the content corresponding to the speaker's gaze point is the same as the content corresponding to the gaze points of each participant, and then confirm whether the participant's device should provide a prompt for the content corresponding to the speaker's gaze point.
[0216] For example, if the content corresponding to the speaker's gaze point overlaps with the content corresponding to the participant's gaze point, the participant is not prompted about the content corresponding to the speaker's gaze point; if the content corresponding to the speaker's gaze point does not overlap with the content corresponding to the participant's gaze point, the participant is prompted about the content corresponding to the speaker's gaze point.
[0217] In some embodiments, if the content corresponding to the speaker's gaze point overlaps with the content corresponding to the participant's gaze point, the participant is not prompted with the content corresponding to the speaker's gaze point; if the content corresponding to the speaker's gaze point does not overlap with the content corresponding to the participant's gaze point, the participant is prompted with the content corresponding to the speaker's gaze point.
[0218] For example, Figure 8A and Figure 8B These are schematic diagrams of a video conferencing scenario provided in the embodiments of this application, and corresponding schematic diagrams of the interfaces of various electronic devices. Figure 8A The scenario shown includes: electronic device A, electronic device B, electronic device C, user A, user B, user C, and server 701. Users A, B, and C are all within the detection range of their respective electronic devices' eye-tracking devices. Electronic devices A, B, and C display the same screen.
[0219] Taking user A as the speaker, and the content indicated by identifier A and the content indicated by identifier C overlap, for example, Figure 8B As shown, electronic device A displays prompt messages 702 and 703 on the interface. Electronic device B displays prompt message 703 on the interface. Electronic device C does not display prompt message 703 on the interface. Prompt messages 701, 702, and 703 can be referred to the corresponding explanations above, and will not be elaborated here.
[0220] In some embodiments, overlap is confirmed by an identifier. The following describes this process in conjunction with... Figure 9 Explain the criteria for determining content overlap.
[0221] For example, Figure 9 This is a flowchart illustrating a content overlap determination method provided in an embodiment of this application. Figure 9 As shown, taking the content indicated by identifier A and the content indicated by identifier B as examples, the method includes:
[0222] S901. Determine the type of content corresponding to identifier A and identifier B.
[0223] In this embodiment, the content does not overlap when the content types are different; when the content types are the same, the corresponding area of the content is determined based on the type. This initial determination of the content type is simple.
[0224] For example, if the content corresponding to both identifier A and identifier B includes controls, S902 can be executed to determine whether the controls overlap; if the content corresponding to both identifier A and identifier B includes text, S903 can be executed to determine whether the text overlaps; if the content corresponding to both identifier A and identifier B includes graphics, S904 can be executed to determine whether the graphics overlap.
[0225] S902. Determine whether the control indicated by label A is the same as the control indicated by label B.
[0226] When the control corresponding to identifier A is the same as the control corresponding to identifier B, the content overlaps (e.g. Figure 10A (As shown); when the control corresponding to label A is different from the control corresponding to label B, the content does not overlap.
[0227] S903. Determine whether the number of overlapping characters in the text corresponding to identifier A and the text corresponding to identifier B is greater than or equal to the overlap threshold A.
[0228] It is understood that identifiers A and B may indicate text in units of characters or in units of paragraphs. When text is indicated in units of characters, overlap can be determined using the above-described S903; when text is indicated in units of paragraphs, overlap can be determined by judging whether the text corresponding to identifier A and the text corresponding to identifier B are located in the same paragraph. This application does not specifically limit the method for determining whether text overlaps in its embodiments.
[0229] The word count of the text can be understood as the number of characters. The overlap threshold A can be 2, 3, or any value; no specific limitation is made here.
[0230] If the number of overlapping characters is greater than or equal to the overlap threshold A, the content overlaps; if the number of overlapping characters is less than the overlap threshold A, the content does not overlap. For example, taking an overlap threshold A of 2 as an example, if... Figure 10B In the image a, gray area 1001 represents the text corresponding to identifier A; gray area 1002 represents the text corresponding to identifier B. The number of overlapping characters is 2, which equals the overlap threshold A, indicating content overlap. If... Figure 10B In the gray area 1003 shown in b, the text corresponding to identifier A is represented; the text corresponding to identifier B is represented in the gray area 1004. The number of overlapping characters is 0, which is less than the overlap threshold A, and the content does not overlap.
[0231] It should be noted that the value of the overlap threshold A is related to the specific application scenario. For example, in scenarios with high precision requirements (e.g., fewer words), the overlap threshold A is smaller; for instance, the overlap threshold A can be 1 (e.g., ...). Figure 10C(as shown in 'a'); for scenarios with lower precision requirements (e.g., lots of text), a larger overlap threshold A is used, for example, an overlap threshold A can be 8 (e.g., ...). Figure 10C (As shown in b in the text). No specific limitations are made here.
[0232] S904. Determine whether the graphic indicated by sign A and the graphic indicated by sign B are the same.
[0233] When the graphic indicated by identifier A and the graphic indicated by identifier B are the same, the content overlaps; when the graphic indicated by identifier A and the graphic indicated by identifier B are different, the content does not overlap. For example, as shown... Figure 10D As shown in Figure 'a', gray area 1005 represents the gaze point region recognized by electronic device A, and gray area 1006 represents the gaze point region recognized by electronic device B. Therefore, the graphic corresponding to label A and the graphic corresponding to label B are both eyes, and thus, the content overlaps. For example... Figure 10D As shown in b, gray area 1007 is the gaze point area recognized by electronic device A, and gray area 1008 is the gaze point area recognized by electronic device B. Then the graphic corresponding to label A is an eye, and the graphic corresponding to label B is a nose. Therefore, the content does not overlap.
[0234] In some embodiments, the messages sent by each electronic device to the server, including an identifier of the content, also carry the system time of each electronic device. The server can then determine whether to send this identifier to other devices based on the system time and the server's current time.
[0235] Specifically, if the difference between the system time and the server's current time is greater than the time threshold A, the server will not send the identifier to other electronic devices; if the difference between the system time and the server's current time is less than or equal to the time threshold A, the server will send the identifier to other electronic devices.
[0236] It is understandable that the time threshold A can be 1 second, 2 seconds, or any value; no specific limitation is made here. In scenarios requiring high precision, the time threshold A can be set to a smaller value, such as in real-time video scenarios where the time threshold A can be set to the millisecond level. In scenarios requiring high accuracy, such as in conferencing scenarios, the time threshold A can be appropriately adjusted to a larger value, such as setting it to the second or minute level.
[0237] In some embodiments, the system time between electronic devices can also be used to determine whether to send the identifier to the electronic device.
[0238] For example, Figure 11 This is a flowchart illustrating a time determination method provided in an embodiment of this application. Taking electronic device A and electronic device B as examples, as follows... Figure 11 As shown, the method includes:
[0239] S1101. Obtain the system time A of electronic device A and the system time B of electronic device B.
[0240] S1102. Convert system time A and system time B to the same time zone.
[0241] S1103. If the absolute value of the difference between system time A and system time B is greater than the time threshold B, the server will not send an identifier from electronic device B to electronic device A; the server will not send an identifier from electronic device A to electronic device B.
[0242] S1104. When the absolute value of the difference between system time A and system time B is less than or equal to the time threshold B, the server sends an identifier from electronic device B to electronic device A; the server sends an identifier from electronic device A to electronic device B.
[0243] Understandably, the time threshold B can be 0 seconds, 1 second, 2 seconds, or any value; no specific limitation is made here. In scenarios requiring high precision, the time threshold B can be set to a smaller value, such as in real-time video scenarios where it can be set to the millisecond level. In scenarios requiring high accuracy, such as in conferencing scenarios, the time threshold B can be appropriately adjusted to a larger value, such as setting it to the second or minute level.
[0244] It is understandable that network lag and other factors can cause message transmission delays, leading to inaccurate content displayed by the server corresponding to the user's gaze. In cases of significant time discrepancies, refraining from sending the identifier of that content to other electronic devices can reduce the likelihood of inaccurate content display.
[0245] Understandably, if multiple users speak in a video conference, each electronic device can provide prompts for the content corresponding to the gaze points of multiple speakers. The specific process is the same as described above. Figure 6 The process shown is similar, and will not be described in detail here.
[0246] Understandably, the server can also confirm whether the content corresponding to the speaker's gaze point is the same as the content corresponding to the gaze points of each participant, and thus confirm whether the speaker should provide prompts for the content corresponding to the gaze points of the participants.
[0247] In some embodiments, if the content corresponding to the participant's gaze point is different from (does not overlap with) the content corresponding to the speaker's gaze point, the electronic device used by the speaker can provide a prompt for that participant. If the content corresponding to the participant's gaze point is the same as the content corresponding to the speaker's gaze point, the electronic device used by the speaker may not provide a prompt for that participant.
[0248] This allows the speaker to highlight attendees who are paying attention to different topics, making it easier for the speaker to identify attendees who are not listening.
[0249] It is understood that the interfaces shown in the above embodiments are illustrated using the example of displaying the same screen. For example, desktop sharing or similar methods can enable various electronic devices to display the same screen. In real-world applications, there are also situations where different electronic devices display different screens. For example, different screens may be displayed on different electronic devices in a two-way video chat, or in virtual meetings or games.
[0250] It should be noted that virtual meetings can simulate real-world meeting scenarios and meeting rooms, creating a visual virtual meeting space and transmitting corresponding images to the participants' devices for display, allowing participants to feel as if they are interacting and communicating in the same space.
[0251] For example, Figure 12A This is a schematic diagram of a virtual meeting scenario provided in an embodiment of this application. Figure 12A The scenario shown includes: electronic device A, electronic device B, electronic device C, user A, user B, user C, and server 1201. Users A, B, and C are all within the detection range of their respective electronic devices' eye-tracking devices. Electronic device A displays an interface 1a generated from user A's perspective; electronic device B displays an interface 1b generated from user B's perspective; and electronic device C displays an interface 1c generated from user C's perspective.
[0252] Interface 1a includes: label 11a corresponding to electronic device A, label 12a corresponding to electronic device B, and label 13a corresponding to electronic device C. In this interface 1a, label 12a is located to the right of label 11a, and label 13a is located to the left of label 11a.
[0253] Interface 1b includes: label 11b corresponding to electronic device A, label 12b corresponding to electronic device B, and label 13b corresponding to electronic device C. In this interface 1b, label 11b is located to the left of label 12b, and label 13b is located to the right of label 12b.
[0254] Interface 1c includes: label 11c corresponding to electronic device A, label 12c corresponding to electronic device B, and label 13c corresponding to electronic device C. In this interface 1c, label 12c is located to the left of label 13c, and label 11c is located to the right of label 13c.
[0255] Electronic device B can transmit identifier B to electronic devices A and C via server 1201. Upon receiving identifier B, electronic device A can display a prompt for identifier 13a on its interface. Upon receiving identifier C, electronic device A can display a prompt for identifier 13c on its interface.
[0256] Electronic device C can transmit identifier C to electronic devices A and B via server 1201. Upon receiving identifier C, electronic device A can display a prompt for identifier 12a on its interface. Electronic device B, upon receiving identifier C, can display a prompt for identifier 12b on its interface. The prompting methods are detailed above and will not be elaborated further here.
[0257] In some embodiments, each electronic device may not display its corresponding user's identifier. For example, identifier 11a may not be displayed in interface 1a; identifier 11b may not be displayed in interface 1b; and identifier 11c may not be displayed in interface 1c.
[0258] For example, such as Figure 12B As shown, electronic device A can display prompt message 1202; electronic device B can display prompt messages 1203 and 1204. Electronic device C can display prompt message 1205; prompt message 1202 is used for indication.
[0259] Message 1202 indicates that users B and C are following each other. Message 1203 indicates that user A is following user C; message 1204 indicates that user C is following you. Message 1205 indicates that both user A and user B are following you.
[0260] In some embodiments, the server can statistically analyze the content corresponding to the user's gaze points on each electronic device to obtain statistical results. The electronic devices can then provide prompts based on these results. The specific implementation is similar to that shown in the embodiments above, and will not be described in detail here.
[0261] It should be noted that the above Figures 9 to 11 In the illustrated embodiment, since the images displayed by each electronic device are identical, the overlap of images can be determined by judging whether the images are the same. When the images displayed by each electronic device are different, the overlap can also be determined by whether the image indicated by marker A and the image indicated by marker B belong to the same object. The object can be a living being such as a person or animal, or a non-living object such as a table, building, or road. No specific limitations are imposed here.
[0262] For example, with Figure 12ATaking the scenario shown as an example, since the graphic corresponding to label 13a and the graphic corresponding to label 13b belong to the same object, the content corresponding to user A's gaze point overlaps with the content corresponding to user B's gaze point. Since the graphic corresponding to label 13a and the graphic corresponding to label 12c do not belong to the same object, the content corresponding to user A's gaze point does not overlap with the content corresponding to user C's gaze point.
[0263] In this way, even when different electronic devices display different images, it can be determined whether users of each device are paying attention to the same object, making it easier to provide prompts based on that object later.
[0264] In some embodiments, each electronic device may provide a notification if other electronic devices are following the user corresponding to that electronic device; otherwise, it may not provide a notification if other electronic devices are following the user corresponding to another electronic device.
[0265] For example, with Figure 12A Taking the scenario shown as an example, electronic device A may not provide any notification; electronic device B may notify user B that user C is paying attention; electronic device C may notify user C that user A and user B are paying attention.
[0266] In real-world applications, mutual following may occur, therefore electronic devices can provide notifications for such situations. The following section explains this phenomenon using mutual following as an example.
[0267] For example, Figure 13 This is a flowchart illustrating a notification method provided in an embodiment of this application. Taking, for example, devices accessing the system via a server, including electronic device 1 and electronic device 2, as shown... Figure 13 As shown, the method includes:
[0268] S1301, Electronic device 1 performs eye tracking to obtain content 1 corresponding to the user 1's gaze point.
[0269] For details, please refer to the above explanations; further details will not be provided here.
[0270] It is understood that electronic device 1 can begin eye tracking after joining a video conference; it can also begin eye tracking after receiving a control indicating to turn on the camera; or it can begin eye tracking after electronic device 1 has acquired an audio signal. This application embodiment does not specifically limit the triggering method for eye tracking.
[0271] S1302. Electronic device A sends message 1 to the server. Message 1 includes: identifier 1 corresponding to content 1.
[0272] Identifier 1 is used to identify content 1. Taking content 1 as an image as an example, identifier 1 can be the name of the image, the name of the component corresponding to the image, etc.
[0273] S1303, Electronic device 1 performs eye tracking to obtain content 1 corresponding to user 1's gaze point.
[0274] S1304. Electronic device 1 sends message 2 to the server. Message 2 includes: identifier 1 corresponding to content 1.
[0275] S1305: If both content 1 and content 2 meet the preset conditions, send message 3 to electronic device 1 and electronic device 2. Message 3 is used to indicate a prompt to follow each other. S1305 includes: S1305-1 and S1305-2. S1305-1: Send message 3 to electronic device 1; S1305-2: Send message 3 to electronic device 1.
[0276] The preset conditions include: the graphic indicated by content 1 originates from electronic device 2, and the graphic indicated by content 2 originates from electronic device 1.
[0277] For example, with Figure 12A In the scenario shown, taking electronic device B as electronic device 1 and electronic device C as electronic device 2 as an example, identifier 13b originates from user C on electronic device C; identifier 12C originates from user B on electronic device B. Therefore, the server can determine that user C and user B are mutually following each other, and thus the server can transmit a message indicating mutual following to electronic devices B and C; upon receiving the message, electronic device B provides a notification regarding the relevant following status. Similarly, electronic device C, upon receiving the message, provides a notification regarding the relevant following status.
[0278] In some embodiments, the preset conditions further include: both the graphic indicated by content 1 and the graphic indicated by content 2 include preset features. The preset features can be features of a face graphic, features of an eye graphic, or features of any graphic, etc., and are not specifically limited here.
[0279] For example, Figure 14A This is a schematic diagram illustrating an application scenario provided by an embodiment of this application. For example... Figure 14A As shown, the scenario includes: electronic device 1, electronic device 2, user 1, user 2, and server 1401. User 1 is within the detection range of the eye-tracking device on electronic device 1. User 2 is within the detection range of the eye-tracking device on electronic device 2.
[0280] When user 1 gazes at the display screen of electronic device 1, electronic device 1 can detect the user 1's gaze point on the display screen through an eye-tracking device, and then obtain the content 1 corresponding to the user 1's gaze point. Electronic device 1 transmits identifier 1 to server 1401, and identifier 1 is used to indicate content 1.
[0281] When user 2 is looking at the display screen of electronic device 2, electronic device 2 can detect the user 2's gaze point on the display screen through an eye-tracking device, and then obtain the content 2 corresponding to the user 2's gaze point. Electronic device 2 transmits identifier 2 to server 1401, and identifier 2 is used to indicate content 2.
[0282] Taking content 1 as the eye of user 2 displayed on electronic device 1, and content 2 as the eye of user 1 displayed on electronic device 2, with the preset feature being an eye graphic, as an example, the server transmits a message indicating mutual attention to electronic devices 1 and 2. After receiving the message indicating mutual attention, electronic devices 1 and 2 provide relevant attention prompts.
[0283] S1306, Electronic device 1 provides a mutual attention prompt.
[0284] S1307, Electronic device 2 provides a mutual attention prompt.
[0285] Mutual attention prompts can include interface animations, hardware vibrations, and sounds. For example, in a two-person video call scenario, "eye contact" can trigger corresponding interface, sound, and / or vibration feedback; in a multi-person meeting scenario, "telepathic connection" can trigger interface animations and / or sound feedback.
[0286] This allows for notifications to be sent to people who follow each other, improving the user experience.
[0287] For example, taking the display of prompts on the interface as an example, such as Figure 12C As shown, electronic devices B and C can display prompt message 1206; prompt message 1206 is used to indicate mutual attention.
[0288] like Figure 14B As shown, electronic device 1 can display the image of user 2, as well as a heart-shaped animation 1402; electronic device 2 can display the image of user 1, as well as a heart-shaped animation 1403.
[0289] Building upon the above embodiments, it is also determined whether there is an overlap in gaze duration. This reduces interference from network lag and improves the accuracy of subsequent prompts.
[0290] Specifically, when the gaze times overlap, the server sends message 3 to both electronic device 1 and electronic device 2. When the gaze times do not overlap, the server does not send message 3 to either electronic device 1 or electronic device 2. The method for determining time overlap can be found above. Figure 11 The corresponding descriptions in the text are not elaborated here.
[0291] In summary, the prompting method provided in this application can be used for eye contact in real-time video communication, such as two-party video calls or multi-party live video calls. When both parties in a video call make eye contact at the same time, interactive feedback can be triggered. It can also be used for real-time meeting sharing and real-time online courses. When multiple parties focus on the same content at the same time, corresponding interactive feedback can be triggered. It can also be used in virtual spaces or gaming fields. When two or more parties focus on the same content at the same time, corresponding interactive feedback can be triggered.
[0292] The prompting method and apparatus provided in this application can be applied to scenarios requiring communication, such as video calls, video conferencing, and gaming.
[0293] It is understood that the above embodiments are illustrated using examples of the server performing content overlap judgment, time overlap judgment, and mutual attention judgment. In some embodiments, the server may not perform these judgments, and each electronic device may perform the corresponding judgments. Specifically, the server may transmit a message to the electronic device indicating the content corresponding to the user's gaze point on other electronic devices. The electronic device may perform the corresponding judgment after receiving the message. The specific judgment method can refer to the server's judgment method described above, and will not be described in detail here.
[0294] In some embodiments, the electronic devices may interact without a server (e.g., electronic device A transmits identifier A via a server). The electronic devices may also have communication connections to interact with each other (e.g., electronic device A may transmit identifier A to electronic device B).
[0295] For example, Figure 15 This is a flowchart illustrating a prompting method provided in an embodiment of this application. Figure 15 As shown, the method includes:
[0296] S1501, the first device and the second device establish a communication connection.
[0297] The communication connection can be of any type, such as any form of wireless communication connection or any form of wired communication connection. No specific limitations are imposed here. Both the first device and the second device can correspond to any of the electronic devices A through C, electronic device 1, and electronic device 2 mentioned above.
[0298] S1502, when the user's gaze content on the first device switches to the first content, the second device provides a prompt for the first content.
[0299] The content of a gaze point can be understood as the content corresponding to that gaze point. The second device can provide prompts in any way, such as voice prompts, display prompts, vibration prompts, etc. No specific limitations are made here.
[0300] In this way, the second device can provide prompts for the content corresponding to the user's gaze point on the first device, making it easier for the user on the second device to understand the content corresponding to the user's gaze point on the first device, thereby improving the interactivity, communication efficiency, accuracy, immersion, and fun of online communication and enhancing the user experience.
[0301] Optionally, the method further includes: the second device displaying a first interface, the first interface including an interaction window between the first device and the second device.
[0302] Taking a meeting scenario as an example, the interactive window can be a window used for conducting meetings, and the first interface can correspond to the above. Figure 5B , Figure 8B , Figure 12B The interface shown; taking a video call scenario as an example, the interaction window can be the video call window. The first interface can correspond to the one described above. Figure 14B The interface shown.
[0303] Optionally, the interactive window displays first content; when the user's gaze content of the first device switches to the first content, a prompt is given for the first content, including: when the user's gaze content of the first device switches to the first content, a first prompt message is displayed on the first interface, the first prompt message is used to indicate that the first content is being paid attention to; and / or, when the user's gaze content of the first device switches to the first content, the first content switches from a first form to a second form, the second form is used to indicate that the first content is being paid attention to, the first form and the second form are different.
[0304] The difference between the first and second forms can be in appearance, such as different sizes, different background colors, and different outline colors.
[0305] This can be achieved by displaying prompts or by changing the format of the primary content. No specific limitations are specified here. Furthermore, it can be applied to scenarios where all content is displayed, such as a shared screen in a meeting.
[0306] Optionally, the interactive window displays second content, and both the second content and the first content belong to the first object; when the user's gaze content of the first device switches to the first content, a prompt is given for the first content, including: when the user's gaze content of the first device switches to the first content, a second prompt message is displayed on the first interface, the second prompt message is used to indicate that the first object is being paid attention to; and / or, when the user's gaze content of the first device switches to the first content, the second content switches from a third form to a fourth form, the fourth form is used to indicate that the first object is being paid attention to, and the third form and the fourth form are different.
[0307] The differences between the third and fourth forms can be in appearance, such as different sizes, different background colors, or different outline colors. The first object can be a living or non-living entity; no specific limitation is made here. The first and second content can be the same or different. For example, the first and second content can correspond to different areas of the same object, or to the same object from different angles, or to different identifiers corresponding to the same object, etc. No specific limitation is made here.
[0308] In this way, prompts can be provided by displaying information or by changing the form of the content belonging to the first object. No specific limitations are made here. Furthermore, this approach can also be applied to scenarios where both elements display the first object, such as the same person in a virtual meeting or the same monster in a game scene.
[0309] Optionally, the interactive window displays content from the first device; when the user's gaze on the first device switches to the first content, and the user's gaze on the second device switches to the third content, and the first content and the third content meet preset conditions, a prompt is given for the first content; the preset conditions include: the first content originates from the second device, and the third content originates from the first device.
[0310] This allows for notifications when users on two devices follow each other, enhancing the interactive experience. For example, a notification can be sent when a user on the first device is looking at the profile picture of a user on the second device, and vice versa. Similarly, in a meeting scenario, a notification can be sent when a user on the first device is looking at a user on the second device, and vice versa.
[0311] Optionally, both the first and third content include features of the preset content.
[0312] The features of the preset content can be human eyes, human faces, or features of any other content. The first and third pieces of content can be the same or different.
[0313] In this way, the user can determine whether to pay attention based on the characteristics of the preset content.
[0314] Optional preset content includes: human eyes and / or human face.
[0315] Optionally, the method further includes prompting for the fourth content when the user's gaze content on the first device switches from the first content to the fourth content.
[0316] In this way, when the gaze point changes, the content of the prompt also changes.
[0317] Optionally, before prompting for the first content, the method further includes: receiving a first message when the user's gaze content of the first device switches to the first content, the first message indicating that a prompt should be made for the first content; prompting for the first content when the user's gaze content of the first device switches to the first content includes: prompting for the first content in response to the first message.
[0318] Optionally, the first message is sent by the server after receiving the second message from the first device, and the second message is used to indicate the first content.
[0319] In this way, the gaze content of the first device can also be obtained from the first device through the server.
[0320] Optionally, the second message carries a first timestamp, which is sent by the server when the difference between the first timestamp and the time when the second message is received is less than a first preset threshold.
[0321] In this way, if the message transmission time is shorter compared to the server's current time, a notification is issued. If the time difference is large, the notification is not sent to other electronic devices, which can reduce the possibility of inaccurate content notifications.
[0322] Optionally, the second message carries a first timestamp, and the method further includes: the second device sending a third message to the server, the third message being used to indicate the gaze content of the user of the second device, the third message carrying a second timestamp; the first message being sent by the server when the absolute value of the difference between the first timestamp and the second timestamp is less than a second preset threshold.
[0323] In this way, if the time difference is large, the identifier of the content will not be sent to other electronic devices, which can reduce the possibility of inaccurate content prompts.
[0324] The above description of the prompting method of the embodiments of this application has been provided. The following description describes the apparatus provided by the embodiments of this application for performing the above method. Those skilled in the art will understand that the methods and apparatus can be combined with and referenced by each other, and the related apparatus provided by the embodiments of this application can perform the steps in the above method.
[0325] like Figure 16 As shown, Figure 16 This is a schematic diagram of a prompting device provided in an embodiment of this application. The interface display device can be an electronic device in the embodiment of this application, or it can be a chip or chip system within an electronic device.
[0326] Figure 16As shown, the prompting device can be used in communication equipment, circuits, hardware components, or chips. The prompting device includes a communication unit 1601 and a processing unit 1602. The communication unit 1601 supports the information interaction steps performed by the prompting method, such as receiving or sending messages; the processing unit 1602 supports the information processing steps performed by the prompting method, such as eye tracking, detecting content changes, and providing prompts.
[0327] When the prompting device is an electronic device, the communication unit 1601 can be a communication interface or an interface circuit. When the prompting device is a chip or chip system within an electronic device, the communication unit 1601 can be a communication interface. For example, the communication interface can be an input / output interface, pins, or circuits, etc.
[0328] In one possible implementation, the interface display device may further include a display unit 1603, which is used to support the prompting device in performing display prompts, such as displaying prompt information or displaying animation effects.
[0329] The prompting devices described in the embodiments of this application may include the units described in the above embodiments.
[0330] Specifically, the processing unit 1602 and the display unit 1603 can be integrated together, and the processing unit 1602 and the display unit 1603 may communicate with each other.
[0331] In one possible implementation, the interface display device may further include a storage unit 1604. The storage unit 1604 may include one or more memories, which may be devices in one or more devices or circuits used to store programs or data.
[0332] The storage unit 1604 can exist independently or be connected to the processing unit 1602 via a communication bus. Alternatively, the storage unit 1604 can be integrated with the processing unit 1602.
[0333] Taking the chip or chip system of the electronic device in the embodiments of this application as an example, the storage unit 1604 can store computer-executable instructions for the methods of the electronic device, so that the processing unit 1602 can execute the methods of the electronic device in the above embodiments. The storage unit 1604 can be a register, cache, or random access memory (RAM), etc., and the storage unit 1604 can be integrated with the processing unit 1602. The storage unit 1604 can be a read-only memory (ROM) or other types of static storage devices that can store static information and instructions, and the storage unit 1604 can be independent of the processing unit 1602.
[0334] The apparatus in this embodiment can be used to execute the steps performed in the above method embodiments, and its implementation principle and technical effect are similar, so they will not be described again here.
[0335] The notification method provided in this application can be applied to electronic devices with communication functions. Electronic devices include terminal devices, and the specific device form of the terminal device can be referred to the above-mentioned descriptions, which will not be repeated here.
[0336] For example, Figure 17 This is a schematic diagram of the structure of an electronic device provided in an embodiment of this application. Figure 17 As shown, the electronic device includes: a processor 110, an external memory interface 120, an internal memory 121, a universal serial bus (USB) interface 130, a charging management module 140, a power management module 141, a battery 142, an antenna 1, an antenna 2, a mobile communication module 150, a wireless communication module 160, an audio module 170, a speaker 170A, a receiver 170B, a microphone 170C, a headphone jack 170D, a sensor module 180, buttons 190, a motor 191, an indicator 192, a camera 193, a display screen 194, and a subscriber identification module (SIM) card interface 195, etc. The sensor module 180 may include a pressure sensor 180A, a gyroscope sensor 180B, a barometric pressure sensor 180C, a magnetic sensor 180D, an accelerometer sensor 180E, a distance sensor 180F, a proximity sensor 180G, a fingerprint sensor 180H, a temperature sensor 180J, a touch sensor 180K, an ambient light sensor 180L, a bone conduction sensor 180M, etc.
[0337] It is understood that the structures illustrated in the embodiments of this application do not constitute a specific limitation on the electronic device. In other embodiments of this application, the electronic device may include more or fewer components than illustrated, or combine some components, or split some components, or have different component arrangements. The illustrated components may be implemented in hardware, software, or a combination of software and hardware.
[0338] The processor 110 may include one or more processing units. These processing units may be independent devices or integrated within one or more processors. The processor 110 may also include a memory for storing instructions and data. For example, the processor 110 may store instructions and data related to a prompting method provided in an embodiment of this application.
[0339] Electronic devices utilize GPUs, displays (194), and application processors to achieve display functions. The GPU is a microprocessor for image processing, connecting the displays (194) and the application processor.
[0340] The display screen 194 is used to display images, videos, etc. The display screen 194 includes a display panel. For example, the display screen 194 can display the aforementioned interface, etc., but this embodiment does not limit the scope of the application.
[0341] Camera 193 can be used to capture images of the user and achieve eye tracking. In addition to camera 193, the electronic device can also be equipped with other eye tracking devices (e.g., eye trackers, virtual reality wearable devices, etc.) or connected to eye tracking devices. No specific limitations are made here.
[0342] The mobile communication module 150 and the wireless communication module 160 are used to enable data interaction between the electronic device and other devices. For example, sending and receiving messages.
[0343] It should be noted that the user information (including but not limited to user device information, user personal information, etc.) and data (including but not limited to data used for analysis, data stored, data displayed, etc.) involved in this application are all information and data authorized by the user or fully authorized by all parties. Furthermore, the collection, use and processing of the relevant data must comply with the relevant laws, regulations and standards of the relevant countries and regions, and corresponding operation portals are provided for users to choose to authorize or refuse.
[0344] This application provides an electronic device, which includes a processor and a memory; the memory stores computer-executable instructions; the processor executes the computer-executable instructions stored in the memory, causing the electronic device to perform the above-described method.
[0345] This application provides a chip. The chip includes a processor, which is used to call a computer program in memory to execute the technical solutions in the above embodiments. Its implementation principle and technical effects are similar to those in the related embodiments described above, and will not be repeated here.
[0346] This application also provides a computer-readable storage medium. The computer-readable storage medium stores a computer program. When the computer program is executed by a processor, it implements the methods described above. The methods described in the above embodiments can be implemented wholly or partially by software, hardware, firmware, or any combination thereof. If implemented in software, the functionality can be stored as one or more instructions or code on or transmitted over the computer-readable medium. The computer-readable medium can include computer storage media and communication media, and can also include any medium that can transfer a computer program from one place to another. The storage medium can be any target medium accessible by a computer.
[0347] In one possible implementation, a computer-readable medium may include random access memory (RAM), read-only memory (ROM), compact discread-only memory (CD-ROM) or other optical disc storage, magnetic disk storage or other magnetic storage devices, or any other medium targeted to carry or to store required program code in the form of instructions or data structures, and accessible by a computer. Furthermore, any connection is appropriately referred to as a computer-readable medium. For example, if software is transmitted from a website, server, or other remote source using coaxial cable, fiber optic cable, twisted pair, digital subscriber line (DSL), or wireless technologies such as infrared, radio, and microwave, then coaxial cable, fiber optic cable, twisted pair, DSL, or wireless technologies such as infrared, radio, and microwave are included in the definition of medium. As used herein, disks and optical discs include optical discs, laser discs, optical discs, digital versatile discs (DVDs), floppy disks, and Blu-ray discs, where disks typically reproduce data magnetically, while optical discs optically reproduce data using lasers. Combinations of the above should also be included within the scope of computer-readable media.
[0348] This application provides a computer program product, which includes a computer program that, when run, causes a computer to perform the above-described method.
[0349] It should be noted that the modules or components shown in the above embodiments can be one or more integrated circuits configured to implement the above methods, such as one or more application-specific integrated circuits (ASICs), one or more digital signal processors (DSPs), or one or more field-programmable gate arrays (FPGAs), etc. Furthermore, when a module is implemented through processing element scheduler code, the processing element can be a general-purpose processor, such as a central processing unit (CPU) or other processors capable of calling program code, such as a controller. Additionally, these modules can be integrated together to implement a system-on-a-chip (SOC).
[0350] In the above embodiments, implementation can be achieved, in whole or in part, through software, hardware, firmware, or any combination thereof. When implemented in software, it can be implemented, in whole or in part, as a computer program product. A computer program product includes one or more computer instructions. When the computer program instructions are loaded and executed on a computer, all or part of the flow or function according to the embodiments of this application is generated. The computer can be a general-purpose computer, a special-purpose computer, a computer network, or other programmable device. The computer instructions can be stored in a computer-readable storage medium or transmitted from one computer-readable storage medium to another. For example, computer instructions can be transmitted from one website, computer, server, or data center to another website, computer, server, or data center via wired (e.g., coaxial cable, fiber optic, digital subscriber line (DSL)) or wireless (e.g., infrared, wireless, microwave, etc.) means. The computer-readable storage medium can be any available medium that a computer can access or a data storage device such as a server or data center that integrates one or more available media. The available medium can be a magnetic medium (e.g., floppy disk, hard disk, magnetic tape), an optical medium (e.g., DVD), or a semiconductor medium (e.g., a solid-state disk (SSD)).
[0351] The term "multiple" in this document refers to two or more. The term "and / or" is merely a description of the relationship between related objects, indicating that three relationships can exist. For example, A and / or B can represent: A alone, A and B simultaneously, or B alone. Furthermore, the character " / " in this document generally indicates an "or" relationship between the preceding and following related objects; in formulas, " / " indicates a "division" relationship. Additionally, it should be understood that in the description of this application, words such as "first" and "second" are used only for descriptive purposes and should not be construed as indicating or implying relative importance or order.
[0352] It should be noted that, in the embodiments of this application, the words "exemplarily" or "for example" are used to indicate examples, illustrations, or explanations. Any embodiment or design scheme described as "exemplarily" or "for example" in this application should not be construed as being more preferred or advantageous than other embodiments or design schemes. Specifically, the use of the words "exemplarily" or "for example" is intended to present the relevant concepts in a specific manner.
[0353] It is understood that the various numerical designations used in the embodiments of this application are merely for descriptive convenience and are not intended to limit the scope of the embodiments of this application.
[0354] It is understood that, in the embodiments of this application, the order of the above-mentioned process numbers does not imply the order of execution. The execution order of each process should be determined by its function and internal logic, and should not constitute any limitation on the implementation process of the embodiments of this application.
Claims
1. A prompting method, characterized in that, Applied to a second device, the method includes: Establish a communication connection with the first device; When the user's gaze on the first device switches to the first content, a prompt is given for the first content.
2. The method according to claim 1, characterized in that, The method further includes: The first interface is displayed, which includes an interaction window between the first device and the second device.
3. The method according to claim 2, characterized in that, The interactive window displays the first content; When the user's gaze on the first device switches to the first content, the prompting for the first content includes: When the user's gaze on the first device switches to the first content, a first prompt message is displayed on the first interface, the first prompt message being used to indicate that the first content is being watched; And / or, when the user's gaze content on the first device switches to the first content, the first content switches from a first form to a second form, the second form being used to indicate that the first content is being watched, and the first form and the second form are different.
4. The method according to claim 2, characterized in that, The interactive window displays second content, and both the second content and the first content belong to the first object; When the user's gaze on the first device switches to the first content, the prompting for the first content includes: When the user's gaze on the first device switches to the first content, a second prompt message is displayed on the first interface, the second prompt message being used to indicate that the first object is being watched; And / or, when the user's gaze content on the first device switches to the first content, the second content switches from the third form to the fourth form, the fourth form being used to indicate that the first object is being focused on, the third form and the fourth form being different.
5. The method according to claim 2, characterized in that, The interactive window displays content from the first device; When the user's gaze content on the first device switches to the first content, the user's gaze content on the second device switches to the third content, and the first content and the third content meet preset conditions, a prompt is given for the first content; The preset conditions include: the first content originates from the second device, and the third content originates from the first device.
6. The method according to claim 5, characterized in that, Both the first content and the third content include the features of preset content.
7. The method according to claim 6, characterized in that, The preset content includes: human eyes and / or human face.
8. The method according to any one of claims 1-7, characterized in that, The method further includes: When the user's gaze on the first device switches from the first content to the fourth content, a prompt is given for the fourth content.
9. The method according to any one of claims 1-8, characterized in that, Before providing the prompt regarding the first content, the method further includes: When the user's gaze on the first device switches to the first content, a first message is received, the first message being used to instruct a prompt for the first content; When the user's gaze on the first device switches to the first content, providing a prompt for the first content includes: responding to the first message by providing a prompt for the first content.
10. The method according to claim 9, characterized in that, The first message is sent by the server after receiving a second message from the first device, and the second message is used to indicate the first content.
11. The method according to claim 10, characterized in that, The second message carries a first timestamp, and the first message is sent by the server when the difference between the first timestamp and the time when the second message is received is less than a first preset threshold.
12. The method according to claim 10 or 11, characterized in that, The second message carries a first timestamp, and the method further includes: A third message is sent to the server, the third message being used to indicate the gaze content of the user of the second device, the third message carrying a second timestamp; The first message is sent by the server when the absolute value of the difference between the first timestamp and the second timestamp is less than a second preset threshold.
13. An electronic device, characterized in that, The electronic device includes: one or more processors and memory; The memory is coupled to the one or more processors, the memory being used to store computer program code, the computer program code including computer instructions, the one or more processors invoking the computer instructions to cause the electronic device to perform the method as described in any one of claims 1 to 12.
14. A chip system, characterized in that, The chip system is applied to an electronic device, the chip system including one or more processors, the one or more processors being used to invoke computer instructions to cause the electronic device to perform the method as described in any one of claims 1 to 12.
15. A computer-readable storage medium, characterized in that, The computer-readable storage medium includes computer instructions that, when executed on an electronic device, cause the electronic device to perform the method as described in any one of claims 1 to 12.
16. A computer program product, characterized in that, The computer program product includes computer program code that, when run on an electronic device, causes the electronic device to perform the method as described in any one of claims 1 to 12.