Interaction method, electronic device, and computer-readable storage medium

By detecting voice messages or sensor information to determine playable voice scenarios, electronic devices can display emoticons associated with the voice and send voice signals when a playable voice scenario is available. This solves the problem of the limited variety of emoticon message presentation and improves the convenience and diversity of interaction.

CN114765641BActive Publication Date: 2026-05-15BEIJING QIHOOD TECHNOLOGY CO LTD
View PDF 3 Cites 0 Cited by

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
BEIJING QIHOOD TECHNOLOGY CO LTD
Filing Date
2021-01-12
Publication Date
2026-05-15

AI Technical Summary

Technical Problem

Existing emoji messages have a limited range of forms and lack the ability to convey attitudes and emotions, thus reducing the convenience of interaction.

Method used

By detecting voice messages or sensor information to determine a playable voice scene, the electronic device displays an emoji pattern associated with the voice while simultaneously sending a voice signal when the voice scene is playable; otherwise, it only displays an emoji pattern.

Benefits of technology

It enriches the interactive forms of emoticons, meets diverse needs, reduces the situation where emoticons containing audio are sent in scenarios where it is not suitable to play audio, and improves user convenience.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN114765641B_ABST
    Figure CN114765641B_ABST
Patent Text Reader

Abstract

The application discloses an interactive method, an electronic device and a computer storage medium. The method comprises the following steps: a first electronic device detects whether it is in a playable voice scene according to one or more of the following: whether a voice message is currently received, or whether specific information is detected through a sensor; when the first electronic device is in the playable voice scene, the first electronic device displays a first expression pattern containing a voice identifier in a first expression interface in response to a first user operation for opening the first expression interface, the first expression pattern being associated with a target voice signal; and the first electronic device sends the first expression pattern and the target voice signal to a second electronic device in response to a second user operation acting on the first expression pattern, so that the convenience can be improved.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] This application relates to the field of electronic technology, specifically to an interaction method, an electronic device, and a computer-readable storage medium. Background Technology

[0002] Currently, electronic devices such as smartphones and tablets are constantly evolving. People can use applications installed on these devices to achieve instant communication, thus providing convenience for communication.

[0003] When exchanging communication messages, users can send text messages, emoticons, or voice messages through instant messaging applications installed on their electronic devices. However, with the increasing prevalence of emoticons, their shortcomings have become apparent: a lack of diverse expression and the inability to convey attitudes and emotions, thus reducing the convenience of interaction. Summary of the Invention

[0004] This application provides an interaction method, an electronic device, and a computer-readable storage medium, which can improve the convenience of interaction.

[0005] In a first aspect, embodiments of this application provide an interaction method, including:

[0006] The first electronic device detects whether it is in a playable voice scenario based on one or more of the following: whether a voice message is currently being received, or whether specific information is detected by a sensor;

[0007] When in the playable voice scenario, in response to a first user operation to open the first emoticon interface, the first electronic device displays a first emoticon pattern containing a voice identifier on the first emoticon interface, and the first emoticon pattern is associated with the target voice signal.

[0008] In response to a second user action on the first emoticon, the first electronic device sends the first emoticon and the target voice signal to the second electronic device, the first emoticon being displayed on the second electronic device and the target voice signal being played on the second electronic device.

[0009] In the interaction method provided in this application embodiment, the first electronic device can determine whether it is in a playable voice scenario by whether it has received a voice message or whether it has detected specific information through a sensor. Voice is sent simultaneously with an emoticon only when the device is in the playable voice scenario. This enriches the interaction forms of emoticons and meets people's increasingly diverse needs. It also reduces the need to send emoticons containing audio in scenarios where playing audio is inconvenient (such as the workplace), thus improving the convenience for users of electronic devices.

[0010] In conjunction with the first aspect of the embodiments of this application, in a first possible implementation of the first aspect of the embodiments of this application, the method further includes: when not in the playable voice scene, in response to a third user operation for opening the second emoticon interface, the first electronic device displays a second emoticon pattern on the second emoticon interface;

[0011] In response to a user action on the second emoticon, the first electronic device sends the second emoticon to the second electronic device.

[0012] In this embodiment, voice is sent simultaneously with emoticons only when the user is in a playable voice scenario. This enriches the interactive forms of emoticons and meets people's increasingly diverse needs. It also reduces the need to send emoticons containing audio in scenarios where playing audio is inconvenient (such as the workplace), thus improving the convenience for users using electronic devices.

[0013] In conjunction with the first aspect of the embodiments of this application, in a second possible implementation of the first aspect of the embodiments of this application, when the first electronic device receives a voice message in a group, the first electronic device determines that it is in the playable voice scenario; wherein, the group contains at least three accounts, and the voice message can be received on the devices corresponding to the at least three accounts;

[0014] The first emoticon interface is the emoticon interface of the group.

[0015] In conjunction with the first aspect of the embodiments of this application, in a third possible implementation of the first aspect of the embodiments of this application, when the first electronic device does not receive a voice message from a set account within a set time, the first electronic device determines that it is not in the playable voice scenario;

[0016] The second emoticon interface is the emoticon interface corresponding to the set account.

[0017] In conjunction with the first aspect of the embodiments of this application, in a fourth possible implementation of the first aspect of the embodiments of this application, when the distance between the location information and the first location information detected by the GPS sensor is less than a first threshold, the first electronic device determines that it is in the playable voice scene;

[0018] When the GPS sensor detects that the distance between the location information and the second location information is less than a second threshold, the first electronic device determines that it is not in the playable voice scene; wherein the first location information and the second location information are different.

[0019] In conjunction with the first aspect of the embodiments of this application, in a fifth possible implementation of the first aspect of the embodiments of this application, in response to a first user operation for opening a first emoticon interface, before the first electronic device displays a first emoticon pattern including a voice identifier on the first emoticon interface, the method further includes:

[0020] In response to a fourth user operation on the first emoticon pattern on the first emoticon interface, the first electronic device calls the microphone to obtain the target voice signal;

[0021] The first electronic device associates the target voice signal with the first facial expression pattern.

[0022] In this embodiment, the target voice signal associated with the first emoticon pattern can also be customized in response to user operation. This enhances the fun and convenience of emoticon interaction and increases the diversity of interaction. Compared to image emoticons which only provide single visual information, the combination of emoticon patterns and voice signals adds auditory information in addition to visual information, thus conveying the sender's attitude and emotions more accurately and vividly, thereby improving interactivity.

[0023] In conjunction with the first aspect of the embodiments of this application, in a sixth possible implementation of the first aspect of the embodiments of this application, the step of the first electronic device sending the first emoticon and the target voice signal to the second electronic device in response to a second user operation acting on the first emoticon includes:

[0024] In response to a second user action on the first emoticon, the first electronic device sends the first emoticon and the target voice signal to the second electronic device, plays the target voice signal, and displays the first emoticon.

[0025] Secondly, embodiments of this application provide a first electronic device, comprising:

[0026] The detection module is used to detect whether a playable voice scene is in operation based on one or more of the following: whether a voice message is currently being received, or whether specific information is detected by a sensor;

[0027] The display module is configured to, when in the playable voice scene, in response to a first user operation to open the first emoticon interface, display a first emoticon pattern containing a voice identifier on the first emoticon interface, wherein the first emoticon pattern is associated with a target voice signal;

[0028] A sending module is configured to send the first emoticon and the target voice signal to a second electronic device in response to a second user operation applied to the first emoticon, wherein the first emoticon is for display on the second electronic device and the target voice signal is for playback on the second electronic device.

[0029] In conjunction with the second aspect of the embodiments of this application, in a first possible implementation of the second aspect of the embodiments of this application, the display module is further configured to, when not in the playable voice scene, in response to a third user operation for opening the second emoticon interface, the first electronic device displays a second emoticon pattern on the second emoticon interface;

[0030] The sending module is further configured to, in response to a user operation applied to the second emoticon pattern, send the second emoticon pattern to the second electronic device.

[0031] In conjunction with the second aspect of the embodiments of this application, in a second possible implementation of the second aspect of the embodiments of this application, the detection module is specifically used to determine that it is in the playable voice scene when a voice message is received in a group; wherein, the group contains at least three accounts, and the voice message can be received on the devices corresponding to the at least three accounts;

[0032] The first emoticon interface is the emoticon interface of the group.

[0033] In conjunction with the second aspect of the embodiments of this application, in a third possible implementation of the second aspect of the embodiments of this application, the detection module is specifically used to determine that it is not in the playable voice scene when no voice message is received from the set account within a set time.

[0034] The second emoticon interface is the emoticon interface corresponding to the set account.

[0035] In conjunction with the second aspect of the embodiments of this application, in a fourth possible implementation of the second aspect of the embodiments of this application, the detection module is specifically used for:

[0036] When the distance between the GPS sensor and the first location information is less than the first threshold, it is determined that the location is in the playable voice scene.

[0037] When the distance between the location information and the second location information detected by the GPS sensor is less than the second threshold, it is determined that the user is not in the playable voice scene; wherein the first location information and the second location information are different.

[0038] In conjunction with the second aspect of the embodiments of this application, in a fifth possible implementation of the second aspect of the embodiments of this application, the first electronic device further includes a microphone module and an association module.

[0039] The microphone module is used to obtain a target voice signal in response to a fourth user operation on the first emoticon pattern applied to the first emoticon interface;

[0040] The association module is used to associate the target voice signal with the first emoticon pattern.

[0041] In conjunction with the second aspect of the embodiments of this application, in a sixth possible implementation of the second aspect of the embodiments of this application, the first electronic device further includes a playback module and a display module:

[0042] The playback module is used to play the target voice signal in response to a second user operation applied to the first emoticon pattern;

[0043] The display module is configured to display the first emoticon in response to a second user operation applied to the first emoticon.

[0044] Thirdly, embodiments of this application provide a first electronic device, which includes: one or more processors and a memory;

[0045] The memory is coupled to the one or more processors, and the memory is used to store computer program code, the computer program code including computer instructions;

[0046] When the one or more processors execute the computer instructions, the first electronic device performs:

[0047] Whether a playable voice scene is detected based on one or more of the following: whether a voice message is currently being received, or whether specific information is detected by sensors;

[0048] When in the playable voice scenario, in response to a first user operation to open the first emoticon interface, a first emoticon pattern containing a voice identifier is displayed on the first emoticon interface, and the first emoticon pattern is associated with the target voice signal.

[0049] In response to a second user action on the first emoticon, the first emoticon and the target voice signal are sent to a second electronic device, the first emoticon being displayed on the second electronic device and the target voice signal being played on the second electronic device.

[0050] In one possible implementation, the processor is further configured to: when not in the playable voice scenario, in response to a third user operation for opening the second emoji interface, display a second emoji pattern on the second emoji interface;

[0051] In response to a user action on the second emoticon, the second emoticon is sent to the second electronic device.

[0052] In one possible implementation, when the first electronic device receives a voice message in a group, the processor determines that it is in the playable voice scenario; wherein the group contains at least three accounts, and the voice message can be received on the devices corresponding to the at least three accounts;

[0053] The first emoticon interface is the emoticon interface of the group.

[0054] In one possible implementation, when the first electronic device does not receive a voice message from a set account within a set time, the processor determines that it is not in the playable voice scenario;

[0055] The second emoticon interface is the emoticon interface corresponding to the set account.

[0056] In one possible implementation, when the distance between the location information and the first location information is less than a first threshold detected by the GPS sensor, the processor determines that it is in the playable voice scene;

[0057] When the distance between the location information and the second location information detected by the GPS sensor is less than a second threshold, the processor determines that it is not in the playable voice scene; wherein the first location information and the second location information are different.

[0058] In one possible implementation, in response to a first user operation to open a first emoji interface, before the first emoji interface displays a first emoji pattern containing a voice identifier, the processor is further configured to perform:

[0059] In response to a fourth user operation on the first emoticon pattern on the first emoticon interface, the microphone is invoked to obtain the target voice signal;

[0060] Associate the target voice signal with the first facial expression pattern.

[0061] In one possible implementation, the processor is further configured to, in response to a second user operation acting on the first emoticon, send the first emoticon and the target voice signal to a second electronic device, play the target voice signal, and display the first emoticon.

[0062] Fourthly, embodiments of this application provide a computer storage medium storing a plurality of instructions, which are adapted to be loaded by a processor and executed by the method steps provided in the first aspect or any implementation thereof of the embodiments of this application.

[0063] Fifthly, embodiments of this application provide a computer program product containing instructions that, when run on a computer or processor, cause the computer or processor to perform the method provided by the first aspect of the embodiments of this application or any possible implementation thereof.

[0064] Understandably, the first electronic device provided in the second aspect, the first electronic device provided in the third aspect, the computer storage medium provided in the fourth aspect, and the computer program product provided in the fifth aspect are all used to execute the interaction method provided in the first aspect. Therefore, the beneficial effects they can achieve can be referred to the beneficial effects in the interaction method provided in the first aspect, and will not be repeated here. Attached Figure Description

[0065] To more clearly illustrate the technical solutions in the embodiments of this application, the accompanying drawings used in the embodiments will be briefly introduced below. Obviously, the drawings described below are only some embodiments of this application. For those skilled in the art, other drawings can be obtained based on these drawings without creative effort.

[0066] Figure 1 This is a flowchart illustrating an interaction method provided in an embodiment of this application;

[0067] Figures 2A to 2D These are some schematic diagrams of interactive interfaces provided in the embodiments of this application;

[0068] Figures 3A-3C These are some schematic diagrams of interactive interfaces provided in the embodiments of this application;

[0069] Figure 4 This is a schematic diagram of the structure of a first electronic device 10 provided in an embodiment of this application;

[0070] Figure 5 This application provides a schematic diagram of the structure of a first electronic device. Detailed Implementation

[0071] In the following description, when referring to the accompanying drawings, the same numbers in different drawings denote the same or similar elements unless otherwise indicated. The embodiments described in the following exemplary embodiments do not represent all embodiments consistent with this application. Rather, they are merely examples of apparatuses and methods consistent with some aspects of this application as detailed in the appended claims.

[0072] In the description of this application, it should be understood that the terms "first," "second," etc., are used for descriptive purposes only and should not be construed as indicating or implying relative importance. Those skilled in the art can understand the specific meaning of the above terms in this application based on the specific circumstances. Furthermore, in the description of this application, unless otherwise stated, "multiple" refers to two or more. "And / or" describes the relationship between related objects, indicating that three relationships can exist. For example, A and / or B can represent: A existing alone, A and B existing simultaneously, or B existing alone. The character " / " generally indicates that the preceding and following related objects are in an "or" relationship.

[0073] To better understand the interaction method and the first electronic device provided in the embodiments of this application, the network architecture used in the embodiments of this application will be described below.

[0074] In this embodiment, the first electronic device can communicate with a network and other devices (such as the second electronic device) via wireless communication technology. This wireless communication technology may include Global System for Mobile Communications (GSM), General Packet Radio Service (GPRS), Code Division Multiple Access (CDMA), Wideband Code Division Multiple Access (WCDMA), Time-Division Code Division Multiple Access (TD-SCDMA), Long Term Evolution (LTE), BT, GNSS, WLAN, NFC, FM, and / or IR technologies. The GNSS may include Global Positioning System (GPS), Global Navigation Satellite System (GLONASS), BeiDou Navigation Satellite System (BDS), Quasi-Zenith Satellite System (QZSS), and / or Satellite Based Augmentation Systems (SBAS).

[0075] Please see Figure 1, Figure 1 This is a flowchart illustrating an interaction method provided in an embodiment of this application. For example... Figure 1 As shown, the interaction method may include steps S101 to S104.

[0076] S101, the first electronic device 10 detects whether it is in a playable voice scene. Specifically, the first electronic device 10 may detect whether it is in a playable voice scene based on one or more of the following: whether a voice message is currently being received, or whether specific information is detected by a sensor.

[0077] S102. When in a playable voice scenario, in response to a first user operation to open the first emoticon interface, the first electronic device 10 displays a first emoticon pattern containing a voice identifier on the first emoticon interface.

[0078] The first emoticon pattern is associated with the target voice signal.

[0079] S103, in response to a second user operation acting on the first emoticon pattern, the first electronic device 10 sends the first emoticon pattern and the target voice signal to the second electronic device 20.

[0080] S104, the second electronic device 20 displays the first emoticon pattern and plays the target voice signal.

[0081] In some embodiments of this application, the first electronic device 10 determines that it is currently in a voice-playable scenario when it detects that it has received a voice message within a recently set time period. The received voice message may include, for example, voice messages from instant messaging applications, or emoticons containing voice within instant messaging applications.

[0082] For examples, please refer to Figures 2A to 2D , Figures 2A to 2D These are schematic diagrams of some interactive interfaces provided in the embodiments of this application. For example... Figure 2A The image shown is a schematic diagram of the chat interface 100 of an instant messaging application (such as WeChat, QQ, etc.) displayed on the first electronic device 10. Figure 2A As shown, the chat interface 100 may include a chat object identifier 101, a return control 102, a text message 103, a voice message 104, a voice control 105, a chat box 106, an emoticon control 107, an expand control 108, and a time identifier 109. Among them:

[0083] The chat participant identifier 101 indicates the identifier of the chat participant. This chat participant can be a single account or a group containing multiple accounts. For example... Figure 2AAs shown, the chat participants can be a group named "Volunteer Association," which can contain up to seven accounts. When a device associated with one of these seven accounts sends a message in the group, all devices associated with the other accounts in the group will receive that message.

[0084] Return control 102 is used to return to the previous level of the chat interface 100.

[0085] For example, a device corresponding to one of the accounts in the group sends a text message 103 within the group. This text message 103 may indicate, "Today I visited a school for migrant workers' children in the suburbs with the volunteer association." Devices corresponding to the other seven accounts (e.g., the first electronic device 10) can then receive this text message 103. Additionally, a device corresponding to one of the accounts in the group can also send a voice message 104 within the group. This voice message 104 may indicate its duration. Devices corresponding to the other seven accounts (e.g., the first electronic device 10) can then also receive this voice message 104. In response to a user operation (e.g., a touch operation) applied to the voice message 104, the first electronic device 10 can play the corresponding voice message. The first electronic device 10 can also convert the voice message 104 into a text message in response to a user operation (e.g., a long press operation).

[0086] The voice control 105 is used to respond to user operations, causing the first electronic device 10 to call the microphone to collect and transmit voice signals. In response to user operations on the voice control 105, the first electronic device 10 can call the microphone to collect and transmit voice signals within the group.

[0087] Chat box 106 is used to receive input text messages, which are then sent out within the group.

[0088] An emoji control 107 is used to cause the first electronic device 10 to unfold and display an emoji interface. In response to a user operation on the emoji control 107, the first electronic device 10 displays an emoji interface, which may contain multiple emoji patterns.

[0089] Expand control 108 to expand more functions, such as "photo function" control, transfer function control, etc.

[0090] The time identifier 109 is used to identify the time when the next message is sent or received. In some embodiments of this application, regarding step S101, when the first electronic device 10 detects that a voice message has been received within the currently set time (e.g., 3 minutes), the first electronic device 10 can determine that it is in a playable voice scene.

[0091] In other embodiments, when the first electronic device receives a voice message in a group, the first electronic device determines that it is in the playable voice scenario; wherein the group contains at least three accounts, and the voice message can be received on the devices corresponding to the at least three accounts; the first emoticon interface is the emoticon interface of the group. For example, when a voice message is detected received within a currently set time (e.g., 3 minutes) and the voice message is received in a group, the first electronic device 10 can determine that it is in the playable voice scenario. For example, as Figure 2A As shown, the current time is 08:08. If voice message 104 is received at 08:07 and it is received in a group, then the first electronic device 10 can be determined to be in a playable voice scene.

[0092] like Figure 2A As shown, regarding step S102, in response to a user operation on the emoji control 107, the first electronic device 10 can display a first emoji interface 200. This first emoji interface 200 may include multiple first emoji patterns containing voice identifiers. These first emoji patterns are associated with a target voice signal. For example... Figure 2B As shown, the first emoticon interface 200 may include multiple first emoticon patterns 201, and each first emoticon pattern may include a voice identifier 202.

[0093] like Figure 2B As shown, in response to a user operation on the first emoticon pattern 201, the first electronic device 10 can send the first emoticon pattern 201 and a target voice signal associated with the first emoticon pattern in the group. The target voice signal can be a voice signal associated with the first emoticon pattern; for example, if the emoticon pattern is a clapping emoticon pattern, then the target voice signal is a clapping voice signal.

[0094] In this embodiment, the devices corresponding to the seven accounts in the group can receive the first emoticon 200 and the target voice signal 202. For example, if the second electronic device 20 is logged into one of the accounts, then the second electronic device 20 can receive the first emoticon 201 and the target voice signal 202. Figure 2C As shown, the second electronic device 20 can display an expression pattern 301 based on the received first expression pattern 201, and play voice based on the target voice signal 202.

[0095] In some embodiments of this application, in response to a second user operation acting on the first emoticon pattern, the first electronic device 10 sends the first emoticon pattern 201 and the target voice signal 202 to the second electronic device 20, plays the target voice signal 202, and displays the first emoticon pattern 201. For example... Figure 2DAs shown, in response to a second user operation on the first emoticon pattern 201, the first electronic device 10 can display the first emoticon pattern 201 on the chat interface 100 and play audio, which is the target voice signal 202.

[0096] In other embodiments of this application, the first electronic device 10 can determine whether it is in the playable voice scene by whether a sensor detects specific information.

[0097] Specifically, in one possible embodiment, when the distance between the location information and the first location information detected by the GPS sensor is less than a first threshold, the first electronic device determines that it is in the playable voice scene. When the distance between the location information and the second location information detected by the GPS sensor is less than a second threshold, the first electronic device determines that it is not in the playable voice scene; wherein the first location information and the second location information are different.

[0098] For example, the first location information could be location information identifying "home" determined by a location application, and the second location information could be location information identifying "company" determined by the location application. Then, when the user is within the "home" area (e.g., the GPS sensor detects that the distance between the location information and the first location information is less than a first threshold), the first electronic device 10 determines that it is in the playable voice scenario. When the user is within the "company" area (e.g., the GPS sensor detects that the distance between the location information and the second location information is less than a second threshold), the first electronic device 10 determines that it is not in the playable voice scenario.

[0099] In some embodiments of this application, the first electronic device 10 may also determine whether it is in a playable voice scenario based on whether it has received a voice message or whether it has detected specific information through a sensor. For example, when the user is within the "home" area (e.g., the GPS sensor detects that the distance between the location information and the first location information is less than a first threshold) and receives a voice message in a group, the first electronic device 10 determines that it is in the playable voice scenario.

[0100] Figure 1 and Figures 2A to 2D In the described embodiment, the first electronic device 10 can determine whether it is in a playable voice scenario by whether it has received a voice message or whether it has detected specific information through a sensor. Voice is sent simultaneously with an emoticon only when the device is in a playable voice scenario. This enriches the interactive forms of emoticons and meets people's increasingly diverse needs. It also reduces the need to send emoticons containing audio in scenarios where playing audio is inconvenient (such as at work), thus improving the convenience for users of electronic devices.

[0101] In some embodiments of this application, when not in the playable voice scenario, in response to a third user operation for opening the second emoticon interface, the first electronic device displays a second emoticon pattern on the second emoticon interface; in response to a user operation applied to the second emoticon pattern, the first electronic device sends the second emoticon pattern to the second electronic device.

[0102] For example, in Figure 2A In the user interface shown, the current time is 08:08. A text message was received at 08:07, and the most recent voice message received in the group was at 06:00. Therefore, the first electronic device 10 can determine that it is not in a playable voice scene. In response to a third user operation to open the second emoticon interface, the first electronic device displays a second emoticon pattern on the second emoticon interface; in response to a user operation on the second emoticon pattern, the first electronic device sends the second emoticon pattern to the second electronic device.

[0103] For example, in a scenario where a chat is being conducted on a first electronic device 10 with a specific account, if the first electronic device does not receive a voice message from the specified account within a set time, the first electronic device determines that it is not in the playable voice scenario; the second emoticon interface is the emoticon interface corresponding to the specified account. The following is in conjunction with... Figures 3A-3C The user interface is shown below. Figures 3A-3C These are some interactive interface diagrams provided in the embodiments of this application.

[0104] like Figure 3A As shown, the chat interface 400, named "Xiao Wang" and corresponding to the chat account, includes an account identifier 401, text messages 402, 403, and 405, time identifiers 404 and 407, and an emoticon control 407. Among them, the account identifier 401 indicates "Xiao Wang".

[0105] The current time can be 08:08. Time markers 404 and 407 indicate that from 11:03 yesterday to 08:07 today, the first electronic device 10 has only interacted with the specific target "Xiao Wang" (corresponding to the designated account) via text messages. That is, the first electronic device 10 did not receive any voice messages from the designated account within the set time (e.g., 10 minutes). Therefore, the first electronic device 10 is determined not to be in the playable voice scenario. The second emoticon interface is the emoticon interface corresponding to this designated account.

[0106] like Figure 3A As shown, in response to a user operation (third user operation) acting on the emoji control 407, the first electronic device 10 displays a second emoji pattern 501 on the second emoji interface 500.

[0107] like Figure 3BAs shown, on the second emoticon interface 500, the second emoticon pattern 501 may no longer include a voice identifier.

[0108] like Figure 3B and Figure 3C As shown, in response to a user operation on the second emoticon pattern 501, the first electronic device 10 sends the second emoticon pattern 501 to the second electronic device 20.

[0109] The second electronic device 20 can display the emoticon pattern 501 according to the received second emoticon pattern 501.

[0110] The first electronic device 10 can also display the emoticon pattern 501.

[0111] In some embodiments of this application, such as Figure 3B As shown, the first electronic device 10 can also display a voice-activated facial expression control 502 on the second facial expression interface 500. This voice-activated facial expression control 502 can be used to cause the first electronic device 10 to display facial expressions containing voice identifiers. In response to user operations on the voice-activated facial expression control 502, the first electronic device 10 can display facial expressions containing voice identifiers on the facial expression interface. (Refer to...) Figures 2B to 2D The described emoticons include voice identifiers.

[0112] In this embodiment, when the first electronic device 10 determines that it is in a playable voice scene, it can display emoticons containing voice identifiers only in the emoticon interface of the chat interface of the set account. For example, the emoticon containing voice identifiers can be displayed only in the emoticon interface of "Xiao Wang's" chat interface. Emoticons without voice identifiers can be displayed in the emoticon interfaces of other chat interfaces.

[0113] In other embodiments of this application, when the first electronic device 10 determines that it is in a playable voice scene, an emoticon pattern containing a voice identifier can be displayed in the emoticon interface of the chat interface of the set account. For example, the emoticon pattern containing the voice identifier can be displayed in the emoticon interface of "Xiao Wang's" chat interface, and the emoticon pattern containing the voice identifier can also be displayed in the emoticon interface of other chat interfaces. This application embodiment does not limit this.

[0114] In some embodiments of this application, the target voice signal associated with the first emoticon pattern may be associated with the first emoticon pattern when the first electronic device 10 downloads the emoticon pattern. In other embodiments of this application, the target voice signal associated with the first electronic device 10 may also be obtained by the first electronic device 10 through a data network according to the meaning of the first emoticon pattern. For example, the first electronic device 10 obtains the voice signal of "clapping" from the data network based on the first emoticon pattern being an "applause" emoticon and associates it with the first emoticon pattern.

[0115] In other embodiments of this application, the target voice signal associated with the first emoticon pattern can also be customized in response to a user operation. Specifically, in response to a first user operation for opening the first emoticon interface, before the first electronic device displays the first emoticon pattern containing a voice identifier on the first emoticon interface, the method further includes: in response to a fourth user operation acting on the first emoticon pattern on the first emoticon interface, the first electronic device calls a microphone to obtain a target voice signal; the first electronic device associates the target voice signal with the first emoticon pattern.

[0116] The fourth user operation could be, for example, a long press. In response to a long press on the first emoji, the first electronic device 10 can prompt "Record the voice corresponding to the emoji," and the first electronic device 10 can activate its microphone to collect the voice signal. The first electronic device associates the target voice signal with the first emoji. Then, subsequently... Figure 2D In the scenario shown, the first electronic device can send the target voice signal and the first emoticon together to the second electronic device 20. The second electronic device 20 can display the first emoticon and play the target voice signal.

[0117] In this embodiment, the voice signal corresponding to the emoticon pattern can be customized by the user, thereby improving the fun and convenience of emoticon interaction and increasing the diversity of interaction. Compared with image emoticons which only have single visual information, the combination of emoticon patterns and voice signals adds auditory information in addition to visual information, so it can more accurately and vividly convey the attitude and emotion of the emoticon sender, thus improving interactivity.

[0118] In other embodiments of this application, after receiving the target voice signal and the first emoticon, if the user opens the corresponding chat interface of the second electronic device 20, the second electronic device 20 may display the first emoticon. The second electronic device 20 may also play the target voice signal only in response to a user operation (e.g., a click) on the first emoticon. This reduces interference caused by automatic playback of the voice signal, ensuring both interactivity and privacy of the voice signal.

[0119] Figure 4 An exemplary schematic diagram of the structure of a first electronic device 10 provided in an embodiment of this application is shown. Figure 4 As shown, the first electronic device 10 may include at least: a detection module 410, a display module 420, and a transmission module 430. Wherein:

[0120] The detection module 410 is used to detect whether a playable voice scene is in accordance with one or more of the following: whether a voice message is currently being received, or whether specific information is detected by a sensor;

[0121] Display module 420 is configured to, when in the playable voice scene, in response to a first user operation to open the first emoticon interface, display a first emoticon pattern containing a voice identifier on the first emoticon interface, wherein the first emoticon pattern is associated with a target voice signal;

[0122] The sending module 430 is configured to send the first emoticon and the target voice signal to a second electronic device in response to a second user operation applied to the first emoticon, wherein the first emoticon is for display on the second electronic device and the target voice signal is for playback on the second electronic device.

[0123] In one possible embodiment, the display module 420 is further configured to, when not in the playable voice scene, in response to a third user operation for opening the second emoticon interface, display a second emoticon pattern on the second emoticon interface;

[0124] The sending module 430 is further configured to, in response to a user operation applied to the second emoticon pattern, send the second emoticon pattern to the second electronic device.

[0125] In one possible embodiment, the detection module 410 is specifically used to determine that the user is in the playable voice scenario when a voice message is received in a group; wherein the group contains at least three accounts, and the voice message can be received on the devices corresponding to the at least three accounts;

[0126] The first emoticon interface is the emoticon interface of the group.

[0127] In one possible embodiment, the detection module 410 is specifically used to determine that it is not in the playable voice scene when no voice message is received from a set account within a set time.

[0128] The second emoticon interface is the emoticon interface corresponding to the set account.

[0129] In one possible embodiment, the detection module 410 is specifically used for:

[0130] When the distance between the GPS sensor and the first location information is less than the first threshold, it is determined that the location is in the playable voice scene.

[0131] When the distance between the location information and the second location information detected by the GPS sensor is less than the second threshold, it is determined that the user is not in the playable voice scene; wherein the first location information and the second location information are different.

[0132] In one possible embodiment, the first electronic device further includes a microphone module and an association module.

[0133] The microphone module is used to obtain a target voice signal in response to a fourth user operation on the first emoticon pattern applied to the first emoticon interface;

[0134] The association module is used to associate the target voice signal with the first emoticon pattern.

[0135] In one possible embodiment, the first electronic device further includes a playback module:

[0136] The playback module is used to play the target voice signal in response to a second user operation applied to the first emoticon pattern;

[0137] The display module is further configured to display the first emoticon in response to a second user operation applied to the first emoticon.

[0138] The first electronic device provided in this application embodiment can determine whether it is in a playable voice scenario by whether a voice message is currently received or whether specific information is detected by a sensor. Voice is sent simultaneously with an emoticon only when the playable voice scenario is in effect. This enriches the interactive forms of emoticons and meets people's increasingly diverse needs. It also reduces the need to send emoticons containing audio in scenarios where audio playback is inconvenient (such as the workplace), thus improving the convenience for users of electronic devices.

[0139] It should be noted that the first electronic device provided in the above embodiments, when executing the interaction method, is only illustrated by the division of the above functional modules. In practical applications, the above functions can be assigned to different functional modules as needed, that is, the internal structure of the device can be divided into different functional modules to complete all or part of the functions described above. In addition, the first electronic device provided in the above embodiments and the interaction method embodiments belong to the same concept, and its implementation process can be found in the method embodiments, which will not be repeated here.

[0140] The sequence numbers of the embodiments in this application are for descriptive purposes only and do not represent the superiority or inferiority of the embodiments.

[0141] Please see Figure 5 This document provides a schematic diagram of the structure of a first electronic device according to an embodiment of this application. Figure 5 As shown, the first electronic device 10 may include: at least one processor 501, at least one network interface 504, user interface 503, memory 505, at least one communication bus 502, GPS sensor 506, and microphone 507.

[0142] The communication bus 502 is used to enable communication between these components.

[0143] The user interface 503 may include a display screen and a camera. Optionally, the user interface 503 may also include a standard wired interface and a wireless interface.

[0144] The network interface 504 may optionally include a standard wired interface or a wireless interface (such as a Wi-Fi interface).

[0145] The processor 501 may include one or more processing cores. The processor 501 connects to various parts of the first electronic device 10 using various interfaces and lines, and performs various functions and processes data of the first electronic device 10 by running or executing instructions, programs, code sets or instruction sets stored in the memory 505, and by calling data stored in the memory 505.

[0146] Optionally, the processor 501 can be implemented using at least one of the following hardware forms: Digital Signal Processing (DSP), Field-Programmable Gate Array (FPGA), and Programmable Logic Array (PLA). The processor 501 can integrate one or a combination of several of the following: Central Processing Unit (CPU), Graphics Processing Unit (GPU), and modem. The CPU primarily handles the operating system, user interface, and applications; the GPU is responsible for rendering and drawing the content required for display; and the modem handles wireless communication. It is understood that the modem can also be implemented as a separate chip without being integrated into the processor 501.

[0147] The memory 505 may include random access memory (RAM) or read-only memory. Optionally, the memory 505 may include a non-transitory computer-readable storage medium. The memory 505 can be used to store instructions, programs, code, code sets, or instruction sets. The memory 505 may include a program storage area and a data storage area, wherein the program storage area may store instructions for implementing an operating system, instructions for at least one function (such as touch function, sound playback function, image playback function, etc.), instructions for implementing the above-described method embodiments, etc.; the data storage area may store data involved in the above-described method embodiments, etc. Optionally, the memory 505 may also be at least one storage device located remotely from the aforementioned processor 501. Figure 5 As shown, the memory 505, which serves as a computer storage medium, may include an operating system, a network communication module, and a user interface module.

[0148] exist Figure 5 In the first electronic device 10 shown, the user interface 503 is mainly used to provide an input interface for the user and to obtain the user's input data; while the processor 501 can be used to call the model training application stored in the memory 505 and specifically perform the following operations:

[0149] Whether a playable voice scene is detected based on one or more of the following: whether a voice message is currently being received, or whether specific information is detected by sensors;

[0150] When in the playable voice scenario, in response to a first user operation to open the first emoticon interface, a first emoticon pattern containing a voice identifier is displayed on the first emoticon interface, and the first emoticon pattern is associated with the target voice signal.

[0151] In response to a second user action on the first emoticon, the first emoticon and the target voice signal are sent to a second electronic device, the first emoticon being displayed on the second electronic device and the target voice signal being played on the second electronic device.

[0152] In one possible embodiment, the processor 501 is further configured to perform: when not in the playable voice scene, in response to a third user operation for opening the second emoticon interface, the first electronic device displays a second emoticon pattern on the second emoticon interface;

[0153] In response to a user action on the second emoticon, the first electronic device sends the second emoticon to the second electronic device.

[0154] In one possible implementation, when the first electronic device receives a voice message in a group, the processor 501 determines that it is in the playable voice scenario; wherein the group contains at least three accounts, and the voice message can be received on the devices corresponding to the at least three accounts;

[0155] The first emoticon interface is the emoticon interface of the group.

[0156] In one possible implementation, when the first electronic device does not receive a voice message from a set account within a set time, the processor 501 determines that it is not in the playable voice scenario;

[0157] The second emoticon interface is the emoticon interface corresponding to the set account.

[0158] GPS sensor 506 is used to obtain positioning information.

[0159] In one possible implementation, when the distance between the location information and the first location information is less than a first threshold detected by the GPS sensor 506, the processor 501 determines that it is in the playable voice scene.

[0160] When the distance between the location information and the second location information is less than the second threshold detected by the GPS sensor 506, the processor 501 determines that it is not in the playable voice scene; wherein the first location information and the second location information are different.

[0161] The first electronic device 10 may also include a microphone 507 for acquiring voice signals.

[0162] In one possible implementation, in response to a first user operation to open a first emoticon interface, before the first emoticon interface displays a first emoticon pattern containing a voice identifier, the processor 501 is further configured to perform:

[0163] In response to a fourth user operation on the first emoticon pattern on the first emoticon interface, the microphone 507 is invoked to obtain the target voice signal;

[0164] Associate the target voice signal with the first facial expression pattern.

[0165] In one possible implementation, the processor 501 is further configured to, in response to a second user operation acting on the first emoticon pattern, send the first emoticon pattern and the target voice signal to a second electronic device, play the target voice signal, and display the first emoticon pattern.

[0166] In this embodiment, the first electronic device can determine whether it is in a playable voice scenario by whether it has received a voice message or whether it has detected specific information through a sensor. Voice is sent simultaneously with an emoticon only when the device is in such a playable voice scenario. This enriches the interactive forms of emoticons and meets people's increasingly diverse needs. It also reduces the need to send emoticons containing audio in scenarios where playing audio is inconvenient (such as at work), thus improving the convenience for users of electronic devices.

[0167] This application also provides a computer-readable storage medium storing instructions that, when executed on a computer or processor, cause the computer or processor to perform the above-described instructions. Figure 1 One or more steps in the illustrated embodiment. If the constituent modules of the first electronic device described above are implemented as software functional units and sold or used as independent products, they can be stored in the computer-readable storage medium.

[0168] In the above embodiments, implementation can be achieved entirely or partially through software, hardware, firmware, or any combination thereof. When implemented using software, it can be implemented entirely or partially as a computer program product. The computer program product includes one or more computer instructions. When the computer program instructions are loaded and executed on a computer, all or part of the processes or functions described in the embodiments of this application are generated. The computer can be a general-purpose computer, a special-purpose computer, a computer network, or other programmable device. The computer instructions can be stored in a computer-readable storage medium or transmitted through the computer-readable storage medium. The computer instructions can be transmitted from one website, computer, server, or data center to another website, computer, server, or data center via wired (e.g., coaxial cable, fiber optic, Digital Subscriber Line (DSL)) or wireless (e.g., infrared, wireless, microwave, etc.) means. The computer-readable storage medium can be any available medium that a computer can access or a data storage device such as a server or data center that integrates one or more available media. The available media may be magnetic media (e.g., floppy disks, hard disks, magnetic tapes), optical media (e.g., digital versatile discs (DVDs)), or semiconductor media (e.g., solid-state drives (SSDs)).

[0169] Those skilled in the art will understand that all or part of the processes in the above embodiments can be implemented by a computer program instructing related hardware. This program can be stored in a computer-readable storage medium, and when executed, it can include the processes of the embodiments of the above methods. The aforementioned storage medium includes various media capable of storing program code, such as Read Only Memory (ROM), Random Access Memory (RAM), magnetic disks, or optical disks. Unless otherwise specified, the technical features of this embodiment and its implementation schemes can be combined arbitrarily.

[0170] The embodiments described above are merely preferred embodiments of this application and are not intended to limit the scope of this application. Any modifications and improvements made by those skilled in the art to the technical solutions of this application without departing from the spirit of this application should fall within the protection scope defined by the claims of this application.

[0171] The interaction method and the first electronic device disclosed in the embodiments of this application have been described in detail above. Specific examples have been used to illustrate the principles and implementation methods of this application. The description of the above embodiments is only for the purpose of helping to understand the method and core ideas of this application. At the same time, for those skilled in the art, there will be changes in the specific implementation methods and application scope based on the ideas of this application. Therefore, the content of this specification should not be construed as a limitation of this application.

Claims

1. An interaction method, characterized in that, include: The first electronic device detects whether it is in a playable voice scene based on one or more of the following: whether a voice message is currently being received, or whether specific information is detected by a sensor, wherein the specific information is positioning information detected by a GPS sensor where the distance between the GPS sensor and the first positioning information corresponding to the playable voice scene is less than a first threshold. When in the playable voice scenario, in response to a first user operation to open the first emoticon interface, the first electronic device displays a first emoticon pattern containing a voice identifier on the first emoticon interface, and the first emoticon pattern is associated with the target voice signal. In response to a second user action on the first emoticon, the first electronic device sends the first emoticon and the target voice signal to the second electronic device, wherein the first emoticon is for display on the second electronic device and the target voice signal is for playback on the second electronic device; The method further includes: When not in the playable voice scenario, in response to a third user operation to open the second emoticon interface, the first electronic device displays a second emoticon pattern on the second emoticon interface, the second emoticon pattern not containing a voice identifier; In response to a user action on the second emoticon, the first electronic device sends the second emoticon to the second electronic device.

2. The method according to claim 1, characterized in that, When the first electronic device receives a voice message in a group, the first electronic device determines that it is in the playable voice scene; wherein, the group contains at least three accounts, and the voice message can be received on the devices corresponding to the at least three accounts; The first emoticon interface is the emoticon interface of the group.

3. The method according to claim 1, characterized in that, If the first electronic device does not receive a voice message from the set account within a set time, the first electronic device determines that it is not in the playable voice scenario; The second emoticon interface is the emoticon interface corresponding to the set account.

4. The method according to claim 1, characterized in that, When the GPS sensor detects that the distance between the location information and the second location information is less than a second threshold, the first electronic device determines that it is not in the playable voice scene; wherein the first location information and the second location information are different.

5. The method according to any one of claims 1-4, characterized in that, In response to a first user operation for opening a first emoji interface, before the first electronic device displays a first emoji pattern including a voice identifier on the first emoji interface, the method further includes: In response to a fourth user operation on the first emoticon pattern on the first emoticon interface, the first electronic device calls the microphone to obtain the target voice signal; The first electronic device associates the target voice signal with the first facial expression pattern.

6. The method according to any one of claims 1-4, characterized in that, In response to a second user action acting on the first emoticon, the first electronic device sends the first emoticon and the target voice signal to the second electronic device, including: In response to a second user action on the first emoticon, the first electronic device sends the first emoticon and the target voice signal to the second electronic device, plays the target voice signal, and displays the first emoticon.

7. A first electronic device, characterized in that, include: The detection module is used to detect whether it is in a playable voice scene based on one or more of the following: whether a voice message is currently received, or whether specific information is detected by a sensor, wherein the specific information is positioning information detected by a GPS sensor where the distance between the GPS sensor and the first positioning information corresponding to the playable voice scene is less than a first threshold. The display module is configured to, when in the playable voice scene, in response to a first user operation to open the first emoticon interface, display a first emoticon pattern containing a voice identifier on the first emoticon interface, wherein the first emoticon pattern is associated with a target voice signal; A sending module is configured to, in response to a second user operation acting on the first emoticon pattern, send the first emoticon pattern and the target voice signal to a second electronic device, wherein the first emoticon pattern is used to be displayed on the second electronic device and the target voice signal is used to be played on the second electronic device. The display module is further configured to, when not in the playable voice scenario, respond to a third user operation for opening the second emoticon interface, the first electronic device displays a second emoticon pattern on the second emoticon interface, the second emoticon pattern not containing a voice identifier; The sending module is further configured to, in response to a user operation applied to the second emoticon pattern, send the second emoticon pattern to the second electronic device.

8. The first electronic device according to claim 7, characterized in that, The detection module is specifically used to determine that the user is in the playable audio scenario when a voice message is received in a group; wherein the group contains at least three accounts, and the voice message can be received on the devices corresponding to the at least three accounts; The first emoticon interface is the emoticon interface of the group.

9. The first electronic device according to claim 7, characterized in that, The detection module is specifically used to determine that the user is not in the playable voice scenario when no voice message is received from a set account within a set time. The second emoticon interface is the emoticon interface corresponding to the set account.

10. The first electronic device according to claim 7, characterized in that, The detection module is also used for: When the distance between the location information and the second location information detected by the GPS sensor is less than the second threshold, it is determined that the user is not in the playable voice scene; wherein the first location information and the second location information are different.

11. The first electronic device according to any one of claims 7-10, characterized in that, The first electronic device also includes a microphone module and an association module. The microphone module is used to obtain a target voice signal in response to a fourth user operation on the first emoticon pattern applied to the first emoticon interface; The association module is used to associate the target voice signal with the first emoticon pattern.

12. The first electronic device according to any one of claims 7-10, characterized in that, The first electronic device also includes a playback module: The playback module is used to play the target voice signal in response to a second user operation applied to the first emoticon pattern; The display module is further configured to display the first emoticon in response to a second user operation applied to the first emoticon.

13. A first electronic device, characterized in that, The first electronic device includes: one or more processors and a memory; The memory is coupled to the one or more processors, and the memory is used to store computer program code, the computer program code including computer instructions; When the one or more processors execute the computer instructions, the first electronic device performs the interaction method as described in any one of claims 1 to 6.

14. A computer-readable storage medium, characterized in that, The instruction includes instructions that, when executed on a first electronic device, cause the first electronic device to perform the interaction method as described in any one of claims 1 to 6.