Message processing method and device, electronic equipment and readable storage medium

By obtaining audio data from the shooting preview interface to generate reply messages, the problem of wasted time due to searching for equipment during shooting is solved, and fast, uninterrupted message reply is achieved.

CN115623321BActive Publication Date: 2026-03-31VIVO MOBILE COMM CO LTD
View PDF 2 Cites 0 Cited by

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2022-10-13
Publication Date
2026-03-31

AI Technical Summary

Technical Problem

During filming, there was a problem of wasted time due to staff searching for equipment to reply to chat messages.

Method used

When a message associated with a user is received in the shooting preview interface, a reply message is generated by obtaining the user's audio data and sent to the target session, avoiding the need for people to search for the device.

Benefits of technology

The shooting process can be completed without interruption, enabling quick message replies and avoiding wasted time due to searching for equipment.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN115623321B_ABST
    Figure CN115623321B_ABST
Patent Text Reader

Abstract

The application discloses a message processing method and device, electronic equipment and a readable storage medium, and belongs to the technical field of communication. The method comprises the following steps: receiving a first message in the case of displaying a shooting preview interface; acquiring audio data corresponding to a first user in the case that the first user is included in the shooting preview interface and is associated with the first message; generating a second message based on the audio data; and sending the second message to a target conversation where the first message is located.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] This application belongs to the field of communication technology, and specifically relates to a message processing method, apparatus, electronic device, and readable storage medium. Background Technology

[0002] Nowadays, it is becoming increasingly common for people to use electronic devices to take pictures. Through these pictures, people can record their lives, work, and other activities anytime and anywhere.

[0003] In one scenario, during filming, if a chat message is received and the person involved in the message happens to be in the filming scene, and the message is important and requires an immediate reply, then after being informed by the relevant person, the person in the filming scene needs to find their device and reply.

[0004] It is evident that, with existing technology, shooting time is wasted because people in the shooting scene are searching for their own equipment. Summary of the Invention

[0005] The purpose of this application embodiment is to provide a message processing method that can solve the problem in the prior art where shooting time is wasted because people in the shooting scene are looking for their own equipment.

[0006] In a first aspect, embodiments of this application provide a message processing method, the method comprising: receiving a first message when a shooting preview interface is displayed; acquiring audio data corresponding to the first user when the shooting preview interface includes a first user associated with the first message; generating a second message based on the audio data; and sending the second message to a target session in which the first message is located.

[0007] Secondly, embodiments of this application provide a message processing apparatus, which includes: a receiving module for receiving a first message when a shooting preview interface is displayed; a first acquisition module for acquiring audio data corresponding to the first user when the shooting preview interface includes a first user associated with the first message; a generating module for generating a second message based on the audio data; and a sending module for sending the second message to the target session where the first message is located.

[0008] Thirdly, embodiments of this application provide an electronic device including a processor and a memory, wherein the memory stores programs or instructions executable on the processor, and the programs or instructions, when executed by the processor, implement the steps of the method described in the first aspect.

[0009] Fourthly, embodiments of this application provide a readable storage medium on which a program or instructions are stored, which, when executed by a processor, implement the steps of the method described in the first aspect.

[0010] Fifthly, embodiments of this application provide a chip, the chip including a processor and a communication interface, the communication interface being coupled to the processor, the processor being used to run programs or instructions to implement the method as described in the first aspect.

[0011] In a sixth aspect, embodiments of this application provide a computer program product stored in a storage medium, which is executed by at least one processor to implement the method described in the first aspect.

[0012] Thus, in the embodiments of this application, when a shooting preview interface is displayed, showing the scene being shot, if a first message is received and is associated with a first user in the shooting preview interface, then taking advantage of the fact that the shooting scene includes the first user, the first user is directly shot to obtain audio data corresponding to the first user. A second message for reply is then generated based on the audio data and sent to the target session where the first message is located. Therefore, based on the embodiments of this application, during the shooting process, the first user can be informed that a reply is required, allowing the first user to directly reply verbally without needing to locate their electronic device, thus avoiding wasted shooting time caused by people in the shooting scene searching for their devices. Attached Figure Description

[0013] Figure 1 This is a flowchart of a message processing method according to an embodiment of this application;

[0014] Figure 2 This is one of the schematic diagrams of the interface of the electronic device according to an embodiment of this application;

[0015] Figure 3 This is a second schematic diagram of the interface of the electronic device according to an embodiment of this application;

[0016] Figure 4 This is the third schematic diagram of the interface of the electronic device according to an embodiment of this application;

[0017] Figure 5 This is the fourth schematic diagram of the interface of the electronic device according to an embodiment of this application;

[0018] Figure 6 This is the fifth schematic diagram of the interface of the electronic device according to an embodiment of this application;

[0019] Figure 7 This is a block diagram of a message processing apparatus according to an embodiment of this application;

[0020] Figure 8 This is one of the hardware structure diagrams of the electronic device according to an embodiment of this application;

[0021] Figure 9 This is the second schematic diagram of the hardware structure of the electronic device according to an embodiment of this application. Detailed Implementation

[0022] The technical solutions of the embodiments of this application will be clearly described below with reference to the accompanying drawings. Obviously, the described embodiments are only some, not all, of the embodiments of this application. All other embodiments obtained by those skilled in the art based on the embodiments of this application are within the scope of protection of this application.

[0023] The terms "first," "second," etc., used in the specification and claims of this application are used to distinguish similar objects and not to describe a specific order or sequence. It should be understood that such use of data can be interchanged where appropriate so that embodiments of this application can be implemented in orders other than those illustrated or described herein, and the objects distinguished by "first," "second," etc., are generally of the same class and the number of objects is not limited; for example, a first object can be one or more. Furthermore, in the specification and claims, "and / or" indicates at least one of the connected objects, and the character " / " generally indicates that the preceding and following objects are in an "or" relationship.

[0024] The message processing method provided in this application embodiment can be executed by a message processing device provided in this application embodiment, or an electronic device integrating the message processing device, wherein the message processing device can be implemented in hardware or software.

[0025] The message processing method provided in this application will be described in detail below with reference to the accompanying drawings, through specific embodiments and application scenarios.

[0026] Figure 1 A flowchart of a message processing method according to an embodiment of this application is shown, exemplified by the method being applied to an electronic device, including:

[0027] Step 110: With the shooting preview interface displayed, receive the first message.

[0028] In this step, in shooting mode, a shooting preview interface is displayed, which is used to display the shooting scene.

[0029] For example, when a user clicks the "Record" button, a shooting preview interface is displayed, showing the shooting scene.

[0030] Optionally, the first message may come from a message from an app other than the camera app.

[0031] Optionally, the first message is a chat message.

[0032] Step 120: If the shooting preview interface includes the first user associated with the first message, obtain the audio data corresponding to the first user.

[0033] In this step, the shooting preview interface includes the first user, that is, the first user is located in the shooting scene, that is, the first user is the person being photographed.

[0034] In this context, the first user is associated with the first message. For example, the first user is mentioned in the first message using the "@" symbol; the first user's name is mentioned in the text content of the first message; or the first user is currently dealing with something mentioned in the first message.

[0035] Optionally, the first user is the native user used for taking the picture.

[0036] Optionally, the first user is not the native user used for taking the picture.

[0037] Optionally, the first message comes from a friend's conversation.

[0038] Optionally, the first message comes from a group session.

[0039] For example, in one scenario, the first user is the local user who is filming and is in charge of filming. After receiving the first message, the first user judges that the message is important and needs to reply immediately. So, while continuing to film, the first user speaks the reply content. In this step, the audio data of the first user replying to the message is obtained.

[0040] For example, in one scenario, the first user is not the local user used for filming, but the local user is in charge of filming. After receiving the first message, the local user judges that the message is important and verbally tells the first user to reply immediately. Then, while continuing filming, the first user speaks the reply, and in this step, the audio data of the first user replying to the message is obtained.

[0041] For example, in one scenario, the first user is the local user who is filming, and a non-local user is in charge of filming. After receiving the first message, the non-local user in charge of filming determines that the message is important, so he verbally tells the first user to reply immediately. Then, while continuing filming, the first user speaks the reply content, and in this step, the audio data of the first user replying to the message is obtained.

[0042] Optionally, the acquisition includes capturing the first user's screen and simultaneously capturing the sound emitted by the first user; the captured sound is then the audio data for this step.

[0043] Step 130: Generate a second message based on the audio data.

[0044] In this step, a second message is generated based on the audio data, serving as a reply to the first message.

[0045] Step 140: Send the second message to the target session where the first message is located.

[0046] In this step, a second message is sent to complete the reply to the first message.

[0047] Thus, in the embodiments of this application, when a shooting preview interface is displayed, showing the scene being shot, if a first message is received and is associated with a first user in the shooting preview interface, then taking advantage of the fact that the shooting scene includes the first user, the first user is directly shot to obtain audio data corresponding to the first user. A second message for reply is then generated based on the audio data and sent to the target session where the first message is located. Therefore, based on the embodiments of this application, during the shooting process, the first user can be informed that a reply is required, allowing the first user to directly reply verbally without needing to locate their electronic device, thus avoiding wasted shooting time caused by people in the shooting scene searching for their devices.

[0048] In the message processing method of another embodiment of this application, step 120 includes:

[0049] Sub-step A1: If the first message includes first user information, determine a first face image that matches the first account image in the shooting preview interface based on the first account image corresponding to the first user information. The first face image corresponds to the first user.

[0050] Optionally, the first user information includes the first user's account name and account number in the chat application.

[0051] For example, the first message uses the "@" function to mention the first user's account name.

[0052] Application scenarios, for example, see Figure 2 During the shooting process, the shooting preview interface 201 is displayed. At this time, the first message 202 is received, in which the user "Zhang San" is "@".

[0053] Furthermore, based on the first user information, the first user's first account image in the chat application is determined.

[0054] In this embodiment, the application scenario is as follows: the first account image includes the face region of the first user.

[0055] Therefore, based on the first account image, it is compared with each face image in the shooting preview interface. If the first face image is found to match the first account image, then the first face image is considered to correspond to the first user.

[0056] Optionally, if a match is successful, a pop-up window is displayed in the vicinity of the first face image in the shooting preview interface. The pop-up window content is used to indicate that the match is successful and to ask the person taking the picture if they agree.

[0057] For example, see Figure 3 The pop-up window 301 displays the prompt: "Do you want to reply to this message?" In addition, pop-up window 301 also includes "OK" and "Cancel" options.

[0058] Optionally, the interpretation of a match includes: the similarity is greater than a certain threshold.

[0059] or,

[0060] Sub-step A2: Receive the first input from the first user in the first message and the shooting preview interface.

[0061] The first input includes touch input performed by the user on the screen, including but not limited to tapping, swiping, and dragging; the first input can also be air input by the user, such as gestures or facial expressions; the first input also includes input performed by the user on physical buttons on the device, including but not limited to pressing. Moreover, the first input includes one or more inputs, which can be continuous or timed.

[0062] In this step, the first input is used to manually match the first user and the first message in the shooting preview interface.

[0063] For example, see Figure 4 The person taking the photo long-presses the first message 401 and drags it to the first user 402 in the shooting preview interface.

[0064] Sub-step A3: In response to the first input, identify the first user in the shooting preview interface.

[0065] Optionally, based on the first input, after matching the first user and the first message in the shooting preview interface, a pop-up window is displayed in the area near the first user in the shooting preview interface. The content of the pop-up window is used to indicate that the match is successful and to ask the person shooting whether they agree.

[0066] In this embodiment, the first user associated with the first message can be determined in the shooting preview interface using both manual and automatic methods, so as to obtain the audio data corresponding to the first user. It is evident that the automatic association method in this embodiment avoids manual operation, while the manual association method meets individual needs and provides accurate association; the two methods are suitable for different scenarios.

[0067] In another embodiment of the message processing method of this application, when a first user associated with the first message is determined in the shooting preview interface, the camera application and the chat application are associated by default so as to directly send the audio data collected by the camera application to the target session of the chat application.

[0068] In another embodiment of the message processing method of this application, three methods are provided for generating a second message.

[0069] Optionally, before generating the second message, a pop-up window is displayed in the shooting preview interface. The content of the pop-up window is used to provide three options corresponding to the three methods in this embodiment, for the photographer to select.

[0070] For example, see Figure 5 After obtaining the audio data, pop-up window 501 is displayed, which shows three options.

[0071] Optionally, based on the pre-set parameters, after acquiring the audio data, a second message can be generated directly according to the pre-set method.

[0072] Step 130 includes at least one of the following:

[0073] Sub-step B1: Determine the audio data as the content of the second message.

[0074] In this step, a first method is provided, which preserves the audio data sent by the first user and generates a voice message.

[0075] Sub-step B2: If the audio data comes from the video data, determine the video content corresponding to the audio data as the content of the second message.

[0076] In this step, a second method is provided: retain the image of the first user sending audio data, and combine it with the audio data sent by the first user to generate a video message.

[0077] Sub-step B3: Determine the text content corresponding to the audio data as the content of the second message.

[0078] In this step, a third method is provided to translate the audio data sent by the first user into text content and generate a text message.

[0079] In this embodiment, the form of the second message is not limited to one type, and at least includes common voice, video and text to enrich the diversity of reply messages.

[0080] In the message processing method of another embodiment of this application, step 120 includes:

[0081] Sub-step C1: Receive the second input, which is used to determine the start time and end time information for acquiring audio data.

[0082] The second input includes touch input performed by the user on the screen, including but not limited to tapping, swiping, and dragging; the second input can also be air input by the user, such as gestures and facial expressions; the second input also includes input by the user on physical buttons on the device, including but not limited to pressing. Moreover, the second input includes one or more inputs, which can be continuous or timed.

[0083] In this application, it is necessary to obtain audio data corresponding to the first user. The audio data is used to generate reply messages and cannot appear in the video content. Therefore, it is necessary to set a target time period in order to obtain the audio data corresponding to the first user within the target time period.

[0084] Optionally, the start time information and end time information correspond to the start time and end time.

[0085] For example, during the shooting process, if the photographer or the first user makes a designated gesture at a certain moment, that moment is determined as the start moment; if the photographer or the first user makes a designated gesture at a later moment, that moment is determined as the end moment.

[0086] Sub-step C2: In response to the second input, acquire the audio data corresponding to the first user within the target time period corresponding to the second input.

[0087] In this step, the time period between the start time and the end time is the target time period.

[0088] Within the target time period, acquire the audio data corresponding to the first user.

[0089] In this embodiment, during the shooting process, the start time and end time information can be determined through a second input, so as to acquire the audio data corresponding to the first user within the target time period between these two time information. Therefore, based on this embodiment, the time period for acquiring the audio data corresponding to the first user is separated from the normal shooting time period. This allows for targeted acquisition of audio data and avoids irrelevant content appearing in the shot video due to the first user's reply message, thereby ensuring shooting quality.

[0090] In another embodiment of the message processing method of this application, the method further includes:

[0091] Step D1: Based on the start time information, obtain the target frame image corresponding to the target time information, where the target time information is the time information preceding the start time information.

[0092] Step D2: Obtain the first image corresponding to the first user in the target frame image.

[0093] Step D3: Based on the shooting preview interface, acquire the first video corresponding to the target time period.

[0094] Step D4: Replace the first user in the first video with the first image.

[0095] In this embodiment, in order to ensure that the shooting is not interrupted, before the shooting video is output after shooting, the target frame image before the start of audio data acquisition is obtained to extract the first image corresponding to the first user in the frame image; and, each frame image of the first video corresponding to the target time period is obtained to replace the first user in each frame image of the first video with the first image.

[0096] Optionally, the target frame image is the frame image corresponding to the time preceding the start time of the target time period.

[0097] Optionally, if the shooting is not finished after the target time period, the frame image corresponding to the next moment after the end of the target time period can be used as the target frame image.

[0098] Optionally, when the first user replies to the message only through verbal expression without adding body language, the replacement in this embodiment is limited to the replacement of the face, which can also preserve the first user's original body language.

[0099] Furthermore, after the replacement is completed, the complete video of this shoot is output.

[0100] In this embodiment, to ensure uninterrupted shooting, a face replacement method can be adopted. In the final output video, the video content generated based on the first user's message reply is adjusted so that the final output video content is coherent, thereby ensuring high video quality.

[0101] In another embodiment of the message processing method of this application, the method further includes:

[0102] Step E1: Based on the shooting preview interface, acquire the second video.

[0103] The second video does not include video data corresponding to the target time period.

[0104] Optionally, in this embodiment, if video data corresponding to the first user needs to be obtained during video recording, the video recording is interrupted.

[0105] For example, start acquiring video data corresponding to the first user and pause recording; further, stop acquiring video data corresponding to the first user and start recording again.

[0106] Furthermore, output the videos taken before and after the pause as the complete video of this shooting session.

[0107] Optionally, the user can manually stitch the two video segments together to create the final video.

[0108] In this embodiment, although the shooting process is interrupted, a way to quickly reply to messages in the shooting scene can still be provided to minimize the time spent replying to messages and avoid wasting time.

[0109] In the message processing method of another embodiment of this application, step 140 includes:

[0110] Sub-step F1: If the user logs into the target session with the second user information and receives the first message, the user sends the second message to the target session where the first message is located based on the second user information.

[0111] In this case, the second user corresponding to the second user information is the same member participating in the target session as the first user; or, the second user corresponding to the second user information is a different member participating in the target session as the first user.

[0112] For example, in one scenario, the first user and the second user are the same user. That is, during the filming process, the person being filmed receives the first message related to themselves and can directly reply to the first message in their own tone.

[0113] Optionally, in this scenario, the target session is a friend's session; or, the target session is a group session.

[0114] For example, in another scenario, the first user and the second user are not the same user. That is, during the filming process, the second user receives the first message, and the first message is related to the first user being filmed. In this case, after obtaining the audio data corresponding to the first user, the second user sends a reply message.

[0115] Optionally, in this scenario, the target session is a group session, and the first user, the second user, and the user who sent the first message are all members participating in the target session.

[0116] In this embodiment, the party receiving the first message can send its own reply to the target session, as well as the reply from others, thereby avoiding delays in filming due to different people searching for their own electronic devices in the filming scene.

[0117] In another embodiment of the message processing method of this application, after step 140, the method further includes:

[0118] Step G1: In the target session interface corresponding to the target session, display the message content of the second message, and display at least one of the following: the third user information corresponding to the first user, the fourth user information that sent the second message, and the referenced first message.

[0119] Optionally, the third user information includes the chat application's user account name, chat application's user account number, and chat application's user account avatar.

[0120] Optionally, the fourth user information includes the chat application's user account name, chat application's user account number, and chat application's user account avatar.

[0121] Optionally, log in to the chat application using a fourth user account.

[0122] For example, in one scenario, if the first user is the local user used for taking photos, and the third and fourth user information are the same, then the message content of the second message, along with the third or fourth user information that sent the second message, will be displayed in the target session. Furthermore, since the second message is a reply to the first message, the reference relationships between the second and fourth messages will also be displayed. For example, the first message will be displayed below the second message in a reference format.

[0123] For example, in one scenario, the first user is not the local user used for taking pictures; both the first user and the local user used for taking pictures are members participating in the target session. Based on this scenario, to clearly describe each message, each user, and the relationship between messages and users, in addition to displaying the message content of the second message, firstly, fourth user information is displayed, i.e., the user information of the user who sent the second message; secondly, third user information is displayed to indicate that this message was replied to by the first user; and thirdly, the referencing relationships existing in the second message are displayed.

[0124] See Figure 6 The second message displayed consists of two parts. One part is displayed in message box 601, including the message content and the phrase "Reply from Zhang San," which indicates the third user information. The other part is displayed below message box 601, including the referenced first message 602. In addition, message box 601 also displays the avatar of the user who sent the message, which indicates the fourth user information.

[0125] In this embodiment, quick replies to messages can be achieved during the shooting process. Since it is not convenient to describe the reply scenario in the target session, in order to clearly describe the reply message, in addition to displaying the message content of the second message, the information of the third user corresponding to the first user, the information of the fourth user who sent the second message, and the referenced first message can also be displayed to explain the relationship between the reply message and the referenced message, and between the reply message and each user, so that each member participating in the target session can understand it at a glance.

[0126] In summary, the purpose of this application is to provide a method for quickly replying to chat messages while filming. First, the filming process can be uninterrupted; second, it provides a new form of interaction, allowing for the acquisition of the subject's audio data for replying during filming; and third, it is simple to operate and provides quick replies.

[0127] The message processing method provided in this application can be executed by a message processing device. This application uses an example of a message processing device executing the message processing method to illustrate the message processing device provided in this application.

[0128] Figure 7 A block diagram of a message processing apparatus according to another embodiment of this application is shown. The apparatus includes:

[0129] The receiving module 10 is used to receive the first message when the shooting preview interface is displayed;

[0130] The first acquisition module 20 is used to acquire audio data corresponding to the first user when the shooting preview interface includes a first user associated with the first message;

[0131] Generation module 30 is used to generate a second message based on audio data;

[0132] The sending module 40 is used to send the second message to the target session where the first message is located.

[0133] Thus, in the embodiments of this application, when a shooting preview interface is displayed, showing the scene being shot, if a first message is received and is associated with a first user in the shooting preview interface, then taking advantage of the fact that the shooting scene includes the first user, the first user is directly shot to obtain audio data corresponding to the first user. A second message for reply is then generated based on the audio data and sent to the target session where the first message is located. Therefore, based on the embodiments of this application, during the shooting process, the first user can be informed that a reply is required, allowing the first user to directly reply verbally without needing to locate their electronic device, thus avoiding wasted shooting time caused by people in the shooting scene searching for their devices.

[0134] Optionally, the first acquisition module 20 includes:

[0135] The first determining unit is configured to, when the first message includes first user information, determine a first face image matching the first account image in the shooting preview interface based on the first account image corresponding to the first user information, wherein the first face image corresponds to the first user.

[0136] The first receiving unit is used to receive the first input from the first user in the first message and the shooting preview interface;

[0137] The second determining unit is used to determine the first user in the shooting preview interface in response to the first input.

[0138] Optionally, the generation module 30 includes:

[0139] The third determining unit is used to determine the audio data as the content of the second message;

[0140] The fourth determining unit is used to determine the video content corresponding to the audio data as the content of the second message when the audio data comes from the video data.

[0141] The fifth determining unit is used to determine the text content corresponding to the audio data as the content of the second message.

[0142] Optionally, the first acquisition module 20 includes:

[0143] The second receiving unit is used to receive a second input, which is used to determine the start time information and end time information for acquiring audio data.

[0144] The acquisition unit is used to acquire audio data corresponding to the first user within a target time period corresponding to the second input in response to the second input.

[0145] Optionally, the device further includes:

[0146] The second acquisition module is used to acquire the target frame image corresponding to the target time information based on the start time information, wherein the target time information is the time information preceding the start time information;

[0147] The third acquisition module is used to acquire the first image corresponding to the first user in the target frame image;

[0148] The fourth acquisition module is used to acquire the first video corresponding to the target time period based on the shooting preview interface;

[0149] The replacement module is used to replace the first user in the first video based on the first image.

[0150] Optionally, the device further includes:

[0151] The fifth acquisition module is used to acquire the third video based on the shooting preview interface;

[0152] The third video does not include video data corresponding to the target time period.

[0153] Optionally, the sending module 40 includes:

[0154] The sending unit is configured to send a second message to the target session where the first message is located based on the second user information when the user logs into the target session with the second user information and receives the first message.

[0155] In this case, the second user corresponding to the second user information is the same member participating in the target session as the first user; or, the second user corresponding to the second user information is a different member participating in the target session as the first user.

[0156] Optionally, the device further includes:

[0157] The display module is used to display the message content of the second message, the third user information corresponding to the first user, the fourth user information that sent the second message, and at least one of the referenced first message in the target session interface corresponding to the target session.

[0158] The message processing device in this application embodiment can be an electronic device or a component within an electronic device, such as an integrated circuit or a chip. The electronic device can be a terminal or other devices besides a terminal. For example, the electronic device can be a mobile phone, tablet computer, laptop computer, PDA, in-vehicle electronic device, mobile internet device (MID), augmented reality (AR) / virtual reality (VR) device, robot, wearable device, ultra-mobile personal computer (UMPC), netbook, or personal digital assistant (PDA), etc. It can also be a server, network attached storage (NAS), personal computer (PC), television set (TV), ATM, or self-service machine, etc. This application embodiment does not specifically limit the device.

[0159] The message processing device in this application embodiment can be a device with an action system. The action system can be an Android action system, an iOS action system, or other possible action systems; this application embodiment does not specifically limit it.

[0160] The message processing apparatus provided in this application embodiment can implement the various processes implemented in the above method embodiments, and will not be described again here to avoid repetition.

[0161] Optionally, such as Figure 8 As shown, this application embodiment also provides an electronic device 100, including a processor 101, a memory 102, and a program or instructions stored in the memory 102 and executable on the processor 101. When the program or instructions are executed by the processor 101, they implement the various steps of any of the above message processing method embodiments and can achieve the same technical effect. To avoid repetition, they will not be described again here.

[0162] It should be noted that the electronic devices in the embodiments of this application include the mobile electronic devices and non-mobile electronic devices described above.

[0163] Figure 9 A schematic diagram of the hardware structure of an electronic device to implement an embodiment of this application.

[0164] The electronic device 1000 includes, but is not limited to, components such as: radio frequency unit 1001, network module 1002, audio output unit 1003, input unit 1004, sensor 1005, display unit 1006, user input unit 1007, interface unit 1008, memory 1009, and processor 1010.

[0165] Those skilled in the art will understand that the electronic device 1000 may also include a power supply (such as a battery) for supplying power to various components. The power supply may be logically connected to the processor 1010 through a power management system, thereby enabling functions such as managing charging, discharging, and power consumption through the power management system. Figure 9 The electronic device structure shown does not constitute a limitation on the electronic device. The electronic device may include more or fewer components than shown, or combine certain components, or have different component arrangements, which will not be elaborated here.

[0166] The processor 1010 is configured to receive a first message when a shooting preview interface is displayed; if the shooting preview interface includes a first user associated with the first message, acquire audio data corresponding to the first user; generate a second message based on the audio data; and send the second message to the target session where the first message is located.

[0167] Thus, in the embodiments of this application, when a shooting preview interface is displayed, showing the scene being shot, if a first message is received and is associated with a first user in the shooting preview interface, then taking advantage of the fact that the shooting scene includes the first user, the first user is directly shot to obtain audio data corresponding to the first user. A second message for reply is then generated based on the audio data and sent to the target session where the first message is located. Therefore, based on the embodiments of this application, during the shooting process, the first user can be informed that a reply is required, allowing the first user to directly reply verbally without needing to locate their electronic device, thus avoiding wasted shooting time caused by people in the shooting scene searching for their devices.

[0168] Optionally, the processor 1010 is further configured to, when the first message includes first user information, determine a first face image matching the first account image in the shooting preview interface based on the first account image corresponding to the first user information, wherein the first face image corresponds to the first user; the user receiving unit 1007 is configured to receive a first input to the first message and the first user in the shooting preview interface; the processor 1010 is further configured to, in response to the first input, determine the first user in the shooting preview interface.

[0169] Optionally, the processor 1010 is further configured to determine the audio data as the content of the second message; if the audio data comes from video data, determine the video content corresponding to the audio data as the content of the second message; and determine the text content corresponding to the audio data as the content of the second message.

[0170] Optionally, the user receiving unit 1007 is further configured to receive a second input, the second input being used to determine start time information and end time information for acquiring the audio data; the processor 1010 is further configured to, in response to the second input, acquire audio data corresponding to the first user within a target time period corresponding to the second input.

[0171] Optionally, the processor 1010 is further configured to: acquire a target frame image corresponding to target time information based on the start time information, wherein the target time information is the time information preceding the start time information; acquire a first image corresponding to the first user in the target frame image; acquire a first video corresponding to the target time period based on the shooting preview interface; and replace the first user in the first video based on the first image.

[0172] Optionally, the processor 1010 is further configured to acquire a third video based on the shooting preview interface; wherein the third video does not include video data corresponding to the target time period.

[0173] Optionally, the processor 1010 is further configured to, upon logging into the target session with the second user information and receiving the first message, send the second message to the target session where the first message is located based on the second user information; wherein the second user corresponding to the second user information and the first user are the same member participating in the target session; or, the second user corresponding to the second user information and the first user are different members participating in the target session.

[0174] Optionally, the display unit 1006 is configured to display the message content of the second message, and at least one of the third user information corresponding to the first user, the fourth user information that sent the second message, and the referenced first message in the target session interface corresponding to the target session.

[0175] In summary, the purpose of this application is to provide a method for quickly replying to chat messages while filming. First, the filming process can be uninterrupted; second, it provides a new form of interaction, allowing for the acquisition of the subject's audio data for replying during filming; and third, it is simple to operate and provides quick replies.

[0176] It should be understood that, in this embodiment, the input unit 1004 may include a graphics processing unit (GPU) 10041 and a microphone 10042. The GPU 10041 processes image data of still images or video images obtained by an image capture device (such as a camera) in video image capture mode or image capture mode. The display unit 1006 may include a display panel 10061, which may be configured in the form of a liquid crystal display, an organic light-emitting diode, etc. The user input unit 1007 includes at least one of a touch panel 10071 and other input devices 10072. The touch panel 10071 is also called a touch screen. The touch panel 10071 may include a touch detection device and a touch controller. Other input devices 10072 may include, but are not limited to, physical keyboards, function keys (such as volume control buttons, power buttons, etc.), trackballs, mice, and joysticks, which will not be described in detail here. The memory 1009 can be used to store software programs and various data, including but not limited to applications and motion systems. Processor 1010 may integrate an application processor and a modem processor. The application processor mainly handles the action system, user page, and applications, while the modem processor mainly handles wireless communication. It is understood that the modem processor may also not be integrated into processor 1010.

[0177] The memory 1009 can be used to store software programs and various data. The memory 1009 may primarily include a first storage area for storing programs or instructions and a second storage area for storing data. The first storage area may store the operating system, application programs or instructions required for at least one function (such as sound playback, image playback, etc.). Furthermore, the memory 1009 may include volatile memory or non-volatile memory, or both. The non-volatile memory may be read-only memory (ROM), programmable read-only memory (PROM), erasable programmable read-only memory (EPROM), electrically erasable programmable read-only memory (EEPROM), or flash memory. Volatile memory can be random access memory (RAM), static random access memory (SRAM), dynamic random access memory (DRAM), synchronous dynamic random access memory (SDRAM), double data rate synchronous dynamic random access memory (DDRSDRAM), enhanced synchronous dynamic random access memory (ESDRAM), synchronous link dynamic random access memory (SLDRAM), and direct memory bus RAM (DRRAM). The memory 1009 in this embodiment includes, but is not limited to, these and any other suitable types of memory.

[0178] The processor 1010 may include one or more processing units; optionally, the processor 1010 integrates an application processor and a modem processor, wherein the application processor mainly handles operations involving the operating system, user interface, and applications, and the modem processor mainly handles wireless communication signals, such as a baseband processor. It is understood that the aforementioned modem processor may also not be integrated into the processor 1010.

[0179] This application also provides a readable storage medium storing a program or instructions. When the program or instructions are executed by a processor, they implement the various processes of the above-described message processing method embodiments and achieve the same technical effects. To avoid repetition, they will not be described again here.

[0180] The processor is the processor in the electronic device described in the above embodiments. The readable storage medium includes computer-readable storage media, such as computer read-only memory (ROM), random access memory (RAM), magnetic disk, or optical disk.

[0181] This application embodiment also provides a chip, which includes a processor and a communication interface. The communication interface is coupled to the processor. The processor is used to run programs or instructions to implement the various processes of the above message processing method embodiments and can achieve the same technical effect. To avoid repetition, it will not be described again here.

[0182] It should be understood that the chip mentioned in the embodiments of this application may also be referred to as a system-on-a-chip, system chip, chip system, or system-on-a-chip, etc.

[0183] This application provides a computer program product that is stored in a storage medium and executed by at least one processor to implement the various processes of the message processing method embodiments described above, and can achieve the same technical effect. To avoid repetition, it will not be described again here.

[0184] It should be noted that, in this document, the terms "comprising," "including," or any other variations thereof are intended to cover non-exclusive inclusion, such that a process, method, article, or apparatus that comprises a list of elements includes not only those elements but also other elements not expressly listed, or elements inherent to such a process, method, article, or apparatus. Without further limitations, an element defined by the phrase "comprising one..." does not exclude the presence of other identical elements in the process, method, article, or apparatus that includes that element. Furthermore, it should be noted that the scope of the methods and apparatuses in the embodiments of this application is not limited to performing functions in the order shown or discussed, but may also include performing functions substantially simultaneously or in the reverse order, depending on the functions involved. For example, the described methods may be performed in a different order than described, and various steps may be added, omitted, or combined. Additionally, features described with reference to certain examples may be combined in other examples.

[0185] Through the above description of the embodiments, those skilled in the art can clearly understand that the methods of the above embodiments can be implemented by means of software plus necessary general-purpose hardware platforms. Of course, they can also be implemented by hardware, but in many cases the former is a better implementation method. Based on this understanding, the technical solution of this application, in essence, or the part that contributes to the prior art, can be embodied in the form of a computer software product. This computer software product is stored in a storage medium (such as ROM / RAM, magnetic disk, optical disk) and includes several instructions to cause a terminal (which may be a mobile phone, computer, server, or network device, etc.) to execute the methods described in the various embodiments of this application.

[0186] The embodiments of this application have been described above with reference to the accompanying drawings. However, this application is not limited to the specific embodiments described above. The specific embodiments described above are merely illustrative and not restrictive. Those skilled in the art can make many other forms under the guidance of this application without departing from the spirit and scope of the claims, and all of these forms are within the protection scope of this application.

Claims

1. A message processing method characterized by, The method is applied to an electronic device, and the method comprises: In a case where a shooting preview interface is displayed, a first message is received; In a case where a shooting object in the shooting preview interface comprises a first user associated with the first message, audio data corresponding to the first user is acquired; the first user is not a native user of the electronic device; Based on the audio data, a second message is generated; the second message is a message replying to the first message; The second message is sent to a target session where the first message is located.

2. The method of claim 1, wherein, The acquisition of the audio data corresponding to the first user in the case where the shooting object in the shooting preview interface comprises the first user associated with the first message comprises: In a case where the first message comprises first user information, a first face image matching a first account image corresponding to the first user is determined from the shooting object in the shooting preview interface according to the first account image corresponding to the first user information; or A first input to the first message and the first user in the shooting preview interface is received; In response to the first input, the first user is determined from the shooting object in the shooting preview interface. The generation of the second message based on the audio data comprises at least one of the following:

3. The method of claim 1, wherein, The audio data is determined as content of the second message; In a case where the audio data is from video data, video content corresponding to the audio data is determined as content of the second message; Text content corresponding to the audio data is determined as content of the second message. The acquisition of the audio data corresponding to the first user comprises:

4. The method of claim 1, wherein, A second input is received, the second input being used to determine start time information and end time information of the acquisition of the audio data; In response to the second input, audio data corresponding to the first user is acquired within a target time period corresponding to the second input. The method further comprises:

5. The method of claim 4, wherein, According to the start time information, a target frame image corresponding to target time information that is a previous time information of the start time information is acquired; A first image corresponding to the first user in the target frame image is acquired; Based on the shooting preview interface, a first video corresponding to the target time period is acquired; According to the first image, the first user in the first video is replaced. The method further comprises:

6. The method of claim 4, wherein, Based on the shooting preview interface, a third video is acquired; In the third video, corresponding video data within the target time period is not included. The sending of the second message to the target session where the first message is located comprises:

7. The method of claim 1, wherein, In a case where a second user logs in the target session with second user information and receives the first message, the second message is sent to the target session where the first message is located based on the second user information; The second user corresponding to the second user information and the first user are the same member participating in the target session, or the second user corresponding to the second user information and the first user are different members participating in the target session. ​ 8. The method of claim 1, wherein, After the sending of the second message to the target session in which the first message is located, the method further includes: displaying, in a target session interface corresponding to the target session, message content of the second message, and at least one of third user information corresponding to the first user, fourth user information sending the second message, and the referenced first message.

9. A message processing device, characterized by The device is applied to an electronic device, and the device includes: a receiving module configured to receive a first message while displaying a shooting preview interface; a first obtaining module configured to, when a shooting object of the shooting preview interface includes a first user associated with the first message, obtain audio data corresponding to the first user, the first user being a non-local user of the electronic device; a generating module configured to generate a second message based on the audio data, the second message being a message replying to the first message; a sending module configured to send the second message to a target session in which the first message is located.

10. An electronic device, comprising: The device includes a processor and a memory, the memory storing programs or instructions executable on the processor, the programs or instructions being executed by the processor to implement the steps of the message processing method according to any one of claims 1-8.

11. A readable storage medium, characterized by, The readable storage medium stores programs or instructions, the programs or instructions being executed by the processor to implement the steps of the message processing method according to any one of claims 1-8.

Citation Information

Patent Citations

  • Communication transferring method, mobile terminal and server

    CN103037319A

  • Shooting method, mobile terminal and computer readable storage medium

    CN107566728A