Information processing method, device, equipment, storage medium and product

By obtaining the business information and user information entered by the user, and using the business speech templates and speech information provided by the server, the business process is automatically recorded and broadcast, which solves the problem of repeated recording in the self-service business processing of financial institutions and improves efficiency and convenience.

CN115550585BActive Publication Date: 2025-09-16CHINA CONSTRUCTION BANK +1
View PDF 3 Cites 0 Cited by

Patent Information

Application Number
CN202211145439.4
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2022-09-20
Publication Date
2025-09-16
Estimated Expiration
2042-09-20

AI Technical Summary

Technical Problem

In the prior art, during the self-service business processing of financial institutions, due to the different readings of the business personnel, repeated recording operations are caused, which reduces the efficiency of business processing.

Method used

By obtaining the business information and user information entered by the user, sending it to the server, receiving the business speech template and speech information, starting the audio and video recording function, and broadcasting the speech nodes in sequence until the recording is completed, the manual reading process is eliminated.

Benefits of technology

It improves the efficiency of business processing, helps users better understand business content, reduces the steps of manual review, and improves operational convenience.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN115550585B_ABST
    Figure CN115550585B_ABST
Patent Text Reader

Abstract

The present application provides an information processing method, device, equipment, storage medium and product, which belongs to the field of data processing technology. The method includes: sending the acquired business information and user information to a server; receiving the business speech template and business speech information sent by the server, and determining the current business speech information according to the user information, business speech template and business speech information, wherein the current business speech information includes at least one speech node; starting the audio and video recording function to record audio and video information, and broadcasting the speech nodes in the current business speech information in sequence until the speech nodes in the current business speech information are broadcast, and the audio and video recording is ended. The method of the present application does not require manual review of the content to be read aloud, and can directly broadcast the current business speech information, eliminating the link of manual reading, facilitating user operation, and effectively improving the efficiency of business handling. Broadcasting in the form of speech nodes facilitates users to better understand the business content.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] The present application relates to the field of data processing technology, and in particular to an information processing method, apparatus, device, storage medium and product. Background Art

[0002] With the continuous advancement of technology, financial institutions usually adopt self-service methods to handle related businesses in order to facilitate customers to handle related businesses. During the process, business personnel read out business-related information to assist customers in handling business.

[0003] During the application process, the entire business process is usually recorded in the form of audio and video. After the recording is completed, it will be reviewed by relevant staff. The relevant staff will review the recorded audio and video to determine whether the business process complies with the relevant procedures.

[0004] However, for different businesses, the content read aloud by business personnel is different. If the content read aloud is incorrect, it needs to be recorded again, which is inconvenient and reduces the efficiency of business handling. Summary of the Invention

[0005] The present application provides an information processing method, apparatus, device, storage medium and product to solve the problem that existing business processing procedures are not convenient enough and the business processing efficiency is low.

[0006] In a first aspect, the present application provides an information processing method, comprising:

[0007] Obtaining business information and user information input by the user based on the business processing interface, and sending the business information and user information to the corresponding server;

[0008] Receive a business speech template and business speech information sent by a server, and determine current business speech information according to the user information, the business speech template and the business speech information, where the current business speech information includes at least one speech node;

[0009] Start the audio and video recording function to record audio and video information, and broadcast the speech nodes in the current business speech information in sequence until all the speech nodes in the current business speech information are broadcast, and end the audio and video recording.

[0010] In a second aspect, the present application provides an information processing device, comprising:

[0011] An acquisition unit, used to acquire business information and user information input by the user based on the business processing interface;

[0012] A transceiver unit, configured to send the service information and the user information to a corresponding server;

[0013] The transceiver unit is also used to receive the business speech template and business speech information sent by the server;

[0014] a determining unit, configured to determine current business speech information according to the user information, the business speech template, and the business speech information, wherein the current business speech information includes at least one speech node;

[0015] The processing unit is used to start the audio and video recording function to record audio and video information, and broadcast the speech nodes in the current business speech information in sequence until all the speech nodes in the current business speech information are broadcast, and end the audio and video recording.

[0016] In a third aspect, the present invention provides an electronic device comprising: a processor, a memory, and a transceiver;

[0017] Processor, memory and transceiver circuit interconnection;

[0018] Memory stores computer-executable instructions;

[0019] a transceiver for sending and receiving information;

[0020] The processor executes the computer-executable instructions stored in the memory, so that the processor performs the method according to the first aspect.

[0021] In a fourth aspect, the present invention provides a computer-readable storage medium, wherein the computer-readable storage medium stores computer-executable instructions, and when the computer-executable instructions are executed by a processor, they are used to implement the method described in the first aspect.

[0022] In a fifth aspect, the present invention provides a computer program product, comprising a computer program, which implements the method described in the first aspect when executed by a processor.

[0023] The information processing method, device, equipment, storage medium and product provided by the present application obtain the business information and user information input by the user based on the business processing interface, and send the business information and the user information to the corresponding server; receive the business speech template and business speech information sent by the server, and determine the current business speech information based on the user information, the business speech template and the business speech information, and the current business speech information includes at least one speech node; start the audio and video recording function to record audio and video information, and broadcast the speech nodes in the current business speech information in sequence until the speech nodes in the current business speech information are broadcast, and end the audio and video recording. There is no need to manually check the content that needs to be read aloud, and the current business speech information can be directly broadcast, eliminating the link of manual reading, facilitating user operation, and effectively improving the efficiency of business processing. In addition, broadcasting in the form of speech nodes facilitates users to better understand the business content. BRIEF DESCRIPTION OF THE DRAWINGS

[0024] The accompanying drawings, which are incorporated in and constitute a part of this specification, illustrate embodiments consistent with the present application and, together with the description, serve to explain the principles of the present application.

[0025] Figure 1 A schematic diagram of the network architecture of the information processing method provided in this application;

[0026] Figure 2 A flowchart of an information processing method provided in this application;

[0027] Figure 3 A flowchart of another information processing method provided in this application;

[0028] Figure 4 A schematic diagram of the structure of an information processing device provided in this application;

[0029] Figure 5 A first block diagram of an electronic device for implementing the information processing method according to an embodiment of the present application;

[0030] Figure 6 This is a second block diagram of an electronic device used to implement the information processing method of an embodiment of the present application.

[0031] The above drawings illustrate specific embodiments of the present application, which will be described in more detail below. These drawings and the textual description are not intended to limit the scope of the present application in any way, but rather to illustrate the concepts of the present application to those skilled in the art by reference to specific embodiments. DETAILED DESCRIPTION

[0032] Exemplary embodiments will be described in detail herein, with examples illustrated in the accompanying drawings. In the following description, when referring to the drawings, identical numerals in different figures represent identical or similar elements, unless otherwise indicated. The embodiments described in the following exemplary embodiments are not intended to represent all embodiments consistent with the present application. Rather, they are merely examples of apparatus and methods consistent with certain aspects of the present application, as detailed in the appended claims.

[0033] In the technical solution of this application, the collection, storage, use, processing, transmission, provision and disclosure of information such as financial data or user data involved comply with the provisions of relevant laws and regulations and do not violate public order and good morals.

[0034] Currently, financial institutions often offer self-service procedures to facilitate customer service. During the process, staff members provide assistance by reading relevant information aloud. This process is often accompanied by dual recording, whereby the entire service process is recorded in both audio and video formats. After the recording is complete, staff review the recordings and review them to determine whether the service was handled in accordance with relevant procedures and whether the customer fully understands the service.

[0035] However, for different businesses, the content read aloud by business personnel is different. If the content read aloud is incorrect, it needs to be recorded again, which is inconvenient and reduces the efficiency of business handling.

[0036] Therefore, in order to solve the problem that the business handling process in the prior art is not convenient enough and the business handling efficiency is low, the inventor found in the research that the business information and user information input by the user based on the business handling interface are obtained, the business information and user information are sent to the server, the business speech template and business speech information sent by the server are received, and the current business speech information is further determined based on the user information, business speech template and business speech information, and the audio and video recording function is started. The speech nodes in the current business speech information are broadcasted in sequence until all the speech nodes are broadcasted, and the audio and video recording is ended. The corresponding business speech information can be obtained for different businesses being handled. There is no need to manually check the content that needs to be read aloud. The current business speech information can be broadcast directly, eliminating the link of manual reading, facilitating user operation, and effectively improving the efficiency of business handling. In addition, the broadcast is carried out in the form of speech nodes, which helps users better understand the business content.

[0037] Therefore, based on the above creative findings, the inventors have proposed the technical solutions of the embodiments of the present application. The network architecture and application scenarios of the information processing method provided by the embodiments of the present application are introduced below.

[0038] like Figure 1As shown, the network architecture corresponding to the public platform operation processing method provided by the embodiment of the present invention includes: a terminal device 1 and a server 2. The terminal device 1 is in communication connection with the server 2, wherein the terminal device 1 has a recording and video recording function. The terminal device 1 obtains the business information and user information input by the user based on the business processing interface, and sends the business information and user information to the corresponding server 2; the server 2 feedbacks the business speech template and business speech information based on the business information and user information; the terminal device 1 receives the business speech template and business speech information sent by the server 2, and the terminal device 1 determines the current business speech information based on the user information, the business speech template and the business speech information. The current business speech information includes at least one speech node; the terminal device 1 starts the audio and video recording function to record audio and video information, and broadcasts the speech nodes in the current business speech information in sequence until the speech nodes in the current business speech information are broadcast, and the audio and video recording ends. The terminal device 1 sends the recorded audio and video information to the server 2, which is stored by the server 2. Different business operations can be handled by providing corresponding business script information, eliminating the need to manually read the content. The current business script information can be directly broadcasted, eliminating the need for manual reading, facilitating user operation and effectively improving business processing efficiency. In addition, the broadcast is in the form of script nodes, which helps users better understand the business content.

[0039] Figure 2 This is a flow chart of an information processing method provided by this application, which is applied to electronic devices such as Figure 2 As shown, the method includes:

[0040] Step 201: Obtain the business information and user information input by the user based on the business processing interface, and send the business information and user information to the corresponding server.

[0041] Electronic devices include servers, terminal devices, wearable devices, etc. Relevant staff use electronic devices to handle services for users. The following uses terminal devices as an example for explanation.

[0042] In this embodiment, the terminal device displays a service processing interface, where the user or relevant staff enters service information and user information. This user information includes the user information of the customer seeking service and / or the user information of the staff assisting the customer in processing the service. The service information includes the service identifier and service name. The user information includes the user name and user identity identifier. The terminal device further obtains the service information and user information entered by the user on the service processing interface and transmits the service information and user information to the corresponding server. The server then responds with the service information and service script information.

[0043] In one possible scenario, a business processing app is pre-installed on the terminal device. The user clicks on the business processing app to display the business processing interface, enters business information and user information, and clicks a confirmation button after completing the input. The terminal device responds to the confirmation operation and obtains the business information and user information entered by the user based on the business processing interface. Alternatively, the user may enter business information and user information using a PC client program or a web browser app or mini-program on the electronic device.

[0044] Optionally, the server obtains the mapping relationship between the preset business information and the business speech template, the server matches the business information sent by the terminal device with the mapping relationship, and the server sends the business speech template corresponding to the preset business information that matches the business information to the terminal device. The server determines the corresponding business speech information based on the user information, such as determining the business tolerance level based on the user information, and the business speech information determined based on the user information at least includes the business tolerance level. The server can also determine the corresponding business speech information based on the business information, and the business information includes the business name and business identifier. The business speech information determined based on the business information includes the name of the branch that handles the business, business details, business type and business level, etc. The above business speech information is combined to obtain the business speech information required by the terminal device.

[0045] Step 202: Receive the business speech template and business speech information sent by the server, and determine the current business speech information according to the user information, the business speech template and the business speech information. The current business speech information includes at least one speech node.

[0046] In this embodiment, the business speech templates and business speech information corresponding to different businesses handled by users will also be different. The terminal device receives the business speech templates and business speech information sent by the server, and determines the speech information for the current business handled by the user, that is, the current business speech information, based on the user information, business speech templates and business speech information. The current business speech information includes at least one speech node, and the speech node refers to the node in the current business speech information that corresponds to the content that may require user confirmation.

[0047] Step 203: Start the audio and video recording function to record audio and video information, and broadcast the speech nodes in the current business speech information in sequence until all the speech nodes in the current business speech information are broadcast, and end the audio and video recording.

[0048] In this embodiment, the terminal device is equipped with a camera and a microphone, and the audio and video recording function is activated to record the audio and video information of the relevant staff and customers conducting business. The speech nodes in the current business speech information are played sequentially, which is equivalent to playing the current business speech information in segments so that the user conducting the business can understand the content. When the speech in the current business speech information is received, the audio and video recording ends.

[0049] In this embodiment, the business information and user information input by the user based on the business processing interface are obtained, the business information and user information are sent to the server, the business speech template and business speech information sent by the server are received, the current business speech information is further determined based on the user information, business speech template and business speech information, the audio and video recording function is started, and the speech nodes in the current business speech information are broadcast in sequence until all the speech nodes are broadcast, and the audio and video recording is ended. The corresponding business speech information can be obtained for different businesses being handled, and there is no need to manually check the content that needs to be read aloud. The current business speech information can be broadcast directly, eliminating the need for manual reading, facilitating user operation, and effectively improving the efficiency of business processing. In addition, the broadcast is carried out in the form of speech nodes, which helps users better understand the business content.

[0050] Figure 3 This is a flow chart of another information processing method provided by this application, which is applied to electronic devices, such as Figure 3 As shown, the method includes:

[0051] Step 301: Obtain the business information and user information input by the user based on the business processing interface, and send the business information and user information to the corresponding server.

[0052] In this embodiment, step 301 has the same technical features as step 201 , and the specific description can refer to step 201 , which will not be repeated here.

[0053] Step 302: Receive the business speech template and business speech information sent by the server, and determine the current business speech information according to the user information, the business speech template and the business speech information. The current business speech information includes at least one speech node.

[0054] In this embodiment, step 302 has the same technical features as step 202 , and the specific description can refer to step 202 , which will not be repeated here.

[0055] In a possible implementation, determining the current business speech information according to the user information, the business speech template, and the business speech information includes:

[0056] Step 2021: Input user information and business speech information into the business speech template to update the business speech template.

[0057] In this embodiment, the information in the business speech template is incomplete and needs to be supplemented with user information and business speech information. The user information and business speech information are input into the business speech template, thereby updating the business speech template.

[0058] Step 2022: Determine whether the information in the updated business speech template is complete.

[0059] In this embodiment, it is determined whether the information in the updated business speech template is complete. Specifically, it is identified whether there is any unfilled information in the updated business speech template; if so, it is determined that the information in the updated business speech template is incomplete; if not, it is determined that the information in the updated business speech template is complete.

[0060] Step 2023: If yes, the updated business speech template is determined as the current business speech information.

[0061] In this embodiment, if the information in the updated business speech template is complete, the updated business speech template is determined as the current business speech information; if the information in the updated business speech template is incomplete, a prompt message of incomplete information is generated and output to prompt the user to complete the missing information.

[0062] Step 303: Start the audio and video recording function to record audio and video information, and broadcast the speech nodes in the current business speech information in sequence until all the speech nodes in the current business speech information are broadcast, and end the audio and video recording.

[0063] In this embodiment, step 303 has the same technical features as step 203. For a detailed description, please refer to step 203 and will not be repeated here.

[0064] In a possible implementation, the speech nodes in the current business speech information are broadcasted in sequence, including:

[0065] Step 3031: After each speech node is broadcast, determine whether there is any content in the speech node that requires user confirmation.

[0066] In this embodiment, after each speech node is broadcast, it is determined whether there is content in the speech node that requires user confirmation, such as whether the customer is clear about the service level.

[0067] Step 3032: If not, continue to broadcast the next speech node.

[0068] In this embodiment, if there is no content in the speech node that requires user confirmation, continue to broadcast the next speech node. After each speech node is broadcast, determine whether there is content in the speech node that requires user confirmation until all the speech nodes in the current business speech information are broadcast, and end the audio and video recording.

[0069] Step 3033: If yes, determine whether the audio and video information contains user confirmation information corresponding to the currently broadcast speech node. If yes, continue to broadcast the next speech node.

[0070] In this embodiment, if a speech node contains content requiring user confirmation, the system determines whether the audio and video information contains user confirmation information corresponding to the currently broadcasted speech node. For example, if the customer is clear about the service level, a user confirmation is required. Further confirmation is then made regarding whether the audio and video information contains user confirmation information corresponding to the currently broadcasted speech node. If user confirmation information exists, the system continues broadcasting the next speech node. After each speech node is broadcast, the system determines whether the speech node contains content requiring user confirmation until all speech nodes in the current service speech information are broadcasted, terminating the audio and video recording. If user confirmation information does not exist, a prompt is generated and output to prompt the user for confirmation.

[0071] In one possible implementation, determining whether the audio and video information contains user confirmation information corresponding to the currently broadcast speech node includes:

[0072] Step 3033a: After each speech node is broadcast, determine the type of speech node currently being broadcast.

[0073] In this embodiment, after each speech node is broadcast, the type of the speech node currently being broadcast is determined. The speech node types are divided into static speech nodes and dynamic speech nodes.

[0074] Step 3033b: If the type of the currently broadcast speech node is a static speech node, determine whether the audio and video information contains user confirmation information corresponding to the currently broadcast speech node based on the recognition result of the audio information in the audio and video information.

[0075] In this embodiment, if the type of the currently broadcast speech node is a static speech node, the user is not required to display the materials required to handle the business, and the audio information in the audio and video needs to be identified. Based on the identification result of the audio information in the audio and video information, it is determined whether the audio and video information contains user confirmation information corresponding to the currently broadcast speech node.

[0076] Optionally, determining whether the audio and video information contains user confirmation information corresponding to the currently broadcast speech node according to the recognition result of the audio information in the audio and video information includes:

[0077] If the recognition result of the audio information contains the preset keywords corresponding to the currently broadcast speech node, it is determined that the audio and video information contains user confirmation information corresponding to the currently broadcast speech node; if the recognition result of the audio information does not contain the preset keywords corresponding to the currently broadcast speech node, it is determined that the audio and video information does not contain user confirmation information corresponding to the currently broadcast speech node.

[0078] In this embodiment, the audio information is recognized, converted into text information, and keywords in the text information are extracted. If the recognition result of the audio information contains the preset keyword corresponding to the currently broadcast speech node, that is, the keyword is a preset keyword, then it is determined that the audio and video information contains user confirmation information corresponding to the currently broadcast speech node. If the recognition result of the audio information does not contain the preset keyword corresponding to the currently broadcast speech node, that is, the keyword is not a preset keyword, then it is determined that the audio and video information does not contain user confirmation information corresponding to the currently broadcast speech node.

[0079] Step 3033c: If the type of the currently broadcast speech node is a dynamic speech node, determine whether the audio and video information contains user confirmation information corresponding to the currently broadcast speech node based on the recognition result of the video information in the audio and video information.

[0080] In this embodiment, if the type of speech node currently being broadcast is a dynamic speech node, the user is required to display the materials required to handle the business, and the video information in the audio and video needs to be identified. Based on the identification result of the video information in the audio and video information, it is determined whether the audio and video information contains user confirmation information corresponding to the currently broadcast speech node.

[0081] Optionally, determining whether the audio and video information contains user confirmation information corresponding to the currently broadcast speech node according to the recognition result of the video information in the audio and video information includes:

[0082] If the recognition result of the video information includes the image information corresponding to the currently played speech node, it is determined that the audio and video information contains user confirmation information corresponding to the currently played speech node; if the recognition result of the video information does not include the image information corresponding to the currently played speech node, it is determined that the audio and video information does not contain user confirmation information corresponding to the currently played speech node.

[0083] In this embodiment, the video information is identified. If the identification result of the video information includes image information corresponding to the currently played speech node, for example, the currently played speech node is to display relevant materials, and the user displays the materials for handling business within the range that the camera can capture, the image information can be identified from the video information, thereby determining that the audio and video information contains user confirmation information corresponding to the currently played speech node; if the identification result of the video information includes image information corresponding to the currently played speech node, it is determined that the audio and video information does not contain user confirmation information corresponding to the currently played speech node, for example, the user may not have displayed the relevant materials, or the display was beyond the shooting range of the camera.

[0084] Step 304: Mark the time node corresponding to the user confirmation information in the audio and video information to obtain the marked audio and video information.

[0085] In this embodiment, in order to facilitate relevant staff to play back audio and video and locate the audio and video, the time node corresponding to the user confirmation information in the audio and video information is marked to obtain the marked audio and video information.

[0086] Step 305: Send the marked audio and video information to the server.

[0087] In this embodiment, the marked audio and video information is sent to the server, and the server stores the marked audio and video information so that relevant staff can retrieve the audio and video information for playback, thereby determining whether the business handling complies with the relevant procedures. By replaying the audio and video information, it can also be understood whether the customer clearly understands the business being handled.

[0088] The current business script information can be directly broadcast without manual reading of the content to be read aloud, eliminating the need for manual reading, facilitating user operation and effectively improving business processing efficiency. Furthermore, the broadcast is presented in the form of script nodes, allowing users to better understand the business content. Furthermore, the time nodes corresponding to user confirmation information in audio and video information are marked, allowing for quick location of audio and video.

[0089] Figure 4 A schematic diagram of the structure of an information processing device provided by this application, such as Figure 4 As shown, the information processing device 400 provided in this embodiment includes an acquiring unit 401 , a transceiver unit 402 , a determining unit 403 , and a processing unit 404 .

[0090] Among them, the acquisition unit 401 is used to obtain the business information and user information input by the user based on the business processing interface. The transceiver unit 402 is used to send the business information and user information to the corresponding server. The transceiver unit 402 is also used to receive the business speech template and business speech information sent by the server. The determination unit 403 is used to determine the current business speech information based on the user information, the business speech template and the business speech information. The current business speech information includes at least one speech node. The processing unit 404 is used to start the audio and video recording function to record audio and video information, and broadcast the speech nodes in the current business speech information in sequence until the speech nodes in the current business speech information are broadcast, and the audio and video recording ends.

[0091] Optionally, the processing unit is also used to determine whether there is content in the speech node that requires user confirmation after each speech node is broadcast; if not, continue to broadcast the next speech node; if so, determine whether the audio and video information contains user confirmation information corresponding to the currently broadcast speech node, and if so, continue to broadcast the next speech node.

[0092] Optionally, the processing unit is also used to determine the type of speech node currently being broadcast after each speech node is broadcast; if the type of the speech node currently being broadcast is a static speech node, then based on the recognition result of the audio information in the audio and video information, determine whether the audio and video information has user confirmation information corresponding to the currently broadcast speech node; if the type of the speech node currently being broadcast is a dynamic speech node, then based on the recognition result of the video information in the audio and video information, determine whether the audio and video information has user confirmation information corresponding to the currently broadcast speech node.

[0093] Optionally, the processing unit is also used to determine that the audio and video information contains user confirmation information corresponding to the currently broadcast speech node if the recognition result of the audio information contains preset keywords corresponding to the currently broadcast speech node; if the recognition result of the audio information does not contain the preset keywords corresponding to the currently broadcast speech node, then determine that the audio and video information does not contain user confirmation information corresponding to the currently broadcast speech node.

[0094] Optionally, the processing unit is also used to determine that the audio and video information contains user confirmation information corresponding to the currently broadcast speech node if the recognition result of the video information contains image information corresponding to the currently broadcast speech node; if the recognition result of the video information does not contain image information corresponding to the currently broadcast speech node, then determine that the audio and video information does not contain user confirmation information corresponding to the currently broadcast speech node.

[0095] Optionally, the determination unit is further used to input user information and business speech information into the business speech template to update the business speech template; determine whether the information in the updated business speech template is complete; if so, determine the updated business speech template as the current business speech information.

[0096] Optionally, the processing unit is further configured to mark the time node corresponding to the user confirmation information in the audio and video information to obtain the marked audio and video information.

[0097] Optionally, the transceiver unit is further configured to send the marked audio and video information to the server.

[0098] Figure 5 A first block diagram of an electronic device for implementing the information processing method according to an embodiment of the present application is shown in FIG. Figure 5 As shown, the electronic device 500 includes: a memory 501 , a processor 502 and a transceiver 503 .

[0099] The processor 502, the memory 501 and the transceiver 503 are interconnected;

[0100] transceiver 503, used for sending and receiving information;

[0101] The memory 501 stores computer-executable instructions;

[0102] The processor 502 executes the computer-executable instructions stored in the memory 501 , so that the processor 502 executes the method provided by any of the above embodiments.

[0103] Figure 6 A second block diagram of an electronic device for implementing the information processing method according to an embodiment of the present application is shown in FIG. Figure 6 As shown, the electronic device can be a computer, a digital broadcast terminal, a messaging device, a tablet device, a personal digital assistant, a server, a server cluster, etc.

[0104] Electronic device 800 may include one or more of the following components: a processing component 802 , a memory 804 , a power component 806 , a multimedia component 808 , an audio component 810 , an input / output (I / O) interface 812 , a sensor component 814 , and a communication component 816 .

[0105] The processing component 802 generally controls the overall operation of the electronic device 800, such as operations associated with display, phone calls, data communications, camera operation, and recording operations. The processing component 802 may include one or more processors 820 to execute instructions to perform all or part of the steps of the above-described method. In addition, the processing component 802 may include one or more modules to facilitate interaction between the processing component 802 and other components. For example, the processing component 802 may include a multimedia module to facilitate interaction between the multimedia component 808 and the processing component 802.

[0106] The memory 804 is configured to store various types of data to support operations on the electronic device 800. Examples of such data include instructions for any application or method operating on the electronic device 800, contact data, phone book data, messages, pictures, videos, etc. The memory 804 can be implemented by any type of volatile or non-volatile storage device, or a combination thereof, such as static random access memory (SRAM), electrically erasable programmable read-only memory (EEPROM), erasable programmable read-only memory (EPROM), programmable read-only memory (PROM), read-only memory (ROM), magnetic memory, flash memory, magnetic disk, or optical disk.

[0107] The power supply component 806 provides power to the various components of the electronic device 800. The power supply component 806 may include a power management system, one or more power supplies, and other components associated with generating, managing, and distributing power to the electronic device 800.

[0108] The multimedia component 808 includes a screen that provides an output interface between the electronic device 800 and the user. In some embodiments, the screen may include a liquid crystal display (LCD) and a touch panel (TP). If the screen includes a touch panel, the screen may be implemented as a touch screen to receive input signals from the user. The touch panel includes one or more touch sensors to sense touches, slides, and gestures on the touch panel. The touch sensor can not only sense the boundaries of a touch or slide action, but also detect the duration and pressure associated with the touch or slide operation. In some embodiments, the multimedia component 808 includes a front camera and / or a rear camera. When the electronic device 800 is in an operating mode, such as a shooting mode or a video mode, the front camera and / or the rear camera can receive external multimedia data. Each front camera and rear camera can be a fixed optical lens system or have a focal length and optical zoom capability.

[0109] The audio component 810 is configured to output and / or input audio signals. For example, the audio component 810 includes a microphone (MIC), which is configured to receive external audio signals when the electronic device 800 is in an operating mode, such as a call mode, a recording mode, and a voice recognition mode. The received audio signal can be further stored in the memory 804 or transmitted via the communication component 816. In some embodiments, the audio component 810 also includes a speaker for outputting audio signals.

[0110] I / O interface 812 provides an interface between processing component 802 and peripheral interface modules, such as a keyboard, click wheel, buttons, etc. These buttons may include but are not limited to: a home button, volume buttons, a start button, and a lock button.

[0111] The sensor assembly 814 includes one or more sensors for providing various aspects of status assessment for the electronic device 800. For example, the sensor assembly 814 can detect the open / closed state of the electronic device 800, the relative positioning of components, such as the display and keypad of the electronic device 800. The sensor assembly 814 can also detect changes in the position of the electronic device 800 or a component of the electronic device 800, the presence or absence of user contact with the electronic device 800, the orientation or acceleration / deceleration of the electronic device 800, and temperature changes of the electronic device 800. The sensor assembly 814 may include a proximity sensor configured to detect the presence of nearby objects without any physical contact. The sensor assembly 814 may also include a light sensor, such as a CMOS or CCD image sensor, for use in imaging applications. In some embodiments, the sensor assembly 814 may also include an accelerometer, a gyroscope sensor, a magnetic sensor, a pressure sensor, or a temperature sensor.

[0112] The communication component 816 is configured to facilitate wired or wireless communication between the electronic device 800 and other devices. The electronic device 800 can access a wireless network based on a communication standard, such as WiFi, 2G or 3G, or a combination thereof. In an exemplary embodiment, the communication component 816 receives a broadcast signal or broadcast-related information from an external broadcast management system via a broadcast channel. In an exemplary embodiment, the communication component 816 also includes a near field communication (NFC) module to facilitate short-range communication. For example, the NFC module can be implemented based on radio frequency identification (RFID) technology, infrared data association (IrDA) technology, ultra-wideband (UWB) technology, Bluetooth (BT) technology and other technologies.

[0113] In an exemplary embodiment, the electronic device 800 may be implemented by one or more application-specific integrated circuits (ASICs), digital signal processors (DSPs), digital signal processing devices (DSPDs), programmable logic devices (PLDs), field programmable gate arrays (FPGAs), controllers, microcontrollers, microprocessors, or other electronic components to perform the above methods.

[0114] In an exemplary embodiment, a non-transitory computer-readable storage medium including instructions is also provided, such as a memory 804 including instructions, and the instructions can be executed by the processor 820 of the electronic device 800 to perform the above method. For example, the non-transitory computer-readable storage medium can be a ROM, a random access memory (RAM), a CD-ROM, a magnetic tape, a floppy disk, an optical data storage device, etc.

[0115] In an exemplary embodiment, a computer-readable storage medium is further provided. The computer-readable storage medium stores computer-executable instructions, and the computer-executable instructions are used by a processor to execute the method in any one of the above embodiments.

[0116] In an exemplary embodiment, a computer program product is further provided, including a computer program. The computer program is used by a processor to execute the method in any one of the above embodiments.

[0117] Those skilled in the art will readily appreciate other embodiments of the present application after considering the specification and practicing the invention disclosed herein. This application is intended to cover any variations, uses, or adaptations of the present application that follow the general principles of the present application and include common knowledge or customary techniques in the art not disclosed herein. The description and examples are to be considered as exemplary only, and the true scope and spirit of the present application are indicated by the following claims.

[0118] It should be understood that the present application is not limited to the exact structure described above and shown in the drawings, and that various modifications and changes may be made without departing from the scope thereof. The scope of the present application is limited only by the appended claims.

Claims

1. An information processing method, characterized in that: The method comprises: Obtaining business information and user information input by the user based on the business processing interface, and sending the business information and user information to the corresponding server; Receive the business speech template and business speech information sent by the server, and determine the current business speech information according to the user information, the business speech template and the business speech information, wherein the current business speech information includes at least one speech node; Start the audio and video recording function to record audio and video information, and broadcast the speech nodes in the current business speech information in sequence until the speech nodes in the current business speech information are broadcast, and end the audio and video recording; The sequentially broadcasting of the speech nodes in the current business speech information includes: After each speech node is broadcast, determine whether there is content in the speech node that requires user confirmation; If not, continue to broadcast the next speech node; If yes, determine whether the audio and video information contains user confirmation information corresponding to the currently broadcast speech node, and if yes, continue to broadcast the next speech node; The step of determining whether the audio and video information contains user confirmation information corresponding to the currently broadcast speech node includes: After each speech node is broadcast, determine the type of the speech node currently being broadcast; If the type of the currently broadcast speech node is a static speech node, determining whether the audio and video information contains user confirmation information corresponding to the currently broadcast speech node based on the recognition result of the audio information in the audio and video information; If the type of the speech node currently being broadcast is a dynamic speech node, then based on the recognition result of the video information in the audio and video information, it is determined whether the audio and video information contains user confirmation information corresponding to the speech node currently being broadcast.

2. The method according to claim 1, characterized in that The determining, based on the recognition result of the audio information in the audio and video information, whether the audio and video information contains user confirmation information corresponding to the currently broadcast speech node includes: If the recognition result of the audio information contains the preset keyword corresponding to the currently broadcast speech node, it is determined that the audio and video information contains user confirmation information corresponding to the currently broadcast speech node; If the recognition result of the audio information does not include the preset keyword corresponding to the currently broadcast speech node, it is determined that the audio and video information does not contain the user confirmation information corresponding to the currently broadcast speech node.

3. The method according to claim 1, characterized in that The determining, based on the recognition result of the video information in the audio and video information, whether the audio and video information includes user confirmation information corresponding to the currently broadcast speech node includes: If the recognition result of the video information includes image information corresponding to the currently played speech node, it is determined that the audio and video information contains user confirmation information corresponding to the currently played speech node; If the recognition result of the video information does not include the image information corresponding to the currently played speech node, it is determined that the audio and video information does not contain the user confirmation information corresponding to the currently played speech node.

4. The method according to any one of claims 1 to 3, characterized in that The determining of the current business speech information according to the user information, the business speech template and the business speech information includes: Inputting the user information and the business speech information into the business speech template to update the business speech template; Determine whether the information in the updated business script template is complete; If so, the updated business speech template is determined as the current business speech information.

5. The method according to claim 1, wherein After the audio and video recording is finished, the method further includes: Mark the time node corresponding to the user confirmation information in the audio and video information to obtain the marked audio and video information.

6. The method according to claim 5, characterized in that After obtaining the marked audio and video information, the method further includes: The marked audio and video information is sent to the server.

7. An information processing device, characterized in that The device comprises: An acquisition unit, used to acquire business information and user information input by the user based on the business processing interface; A transceiver unit, configured to send the service information and the user information to a corresponding server; The transceiver unit is also used to receive the business speech template and business speech information sent by the server; a determining unit, configured to determine current business speech information according to the user information, the business speech template, and the business speech information, wherein the current business speech information includes at least one speech node; A processing unit is configured to start an audio and video recording function to record audio and video information, and to broadcast the speech nodes in the current business speech information in sequence until all the speech nodes in the current business speech information are broadcast, thereby ending the audio and video recording; The processing unit is further configured to, after each speech node is broadcast, determine whether there is content in the speech node that requires user confirmation; if not, continue to broadcast the next speech node; if so, determine whether the audio and video information contains user confirmation information corresponding to the currently broadcast speech node; if so, continue to broadcast the next speech node; The processing unit is further used to determine the type of speech node currently being broadcast after each speech node is broadcast; if the type of the speech node currently being broadcast is a static speech node, then based on the recognition result of the audio information in the audio and video information, determine whether the audio and video information has user confirmation information corresponding to the currently broadcast speech node; if the type of the speech node currently being broadcast is a dynamic speech node, then based on the recognition result of the video information in the audio and video information, determine whether the audio and video information has user confirmation information corresponding to the currently broadcast speech node.

8. An electronic device comprising: processors, memory, and transceivers; Processor, memory and transceiver circuit interconnection; a transceiver for sending and receiving information; The memory stores computer-executable instructions; The processor executes the computer-executable instructions stored in the memory, so that the processor performs the method according to any one of claims 1 to 6.

9. A computer-readable storage medium, characterized in that The computer-readable storage medium stores computer-executable instructions, which are used to implement the method according to any one of claims 1 to 6 when executed by a processor.

10. A computer program product comprising a computer program, wherein when the computer program is executed by a processor, the method according to any one of claims 1 to 6 is implemented.

Citation Information

Patent Citations

  • Video recording method and device, computer equipment and storage medium

    CN110266981A

  • Business data processing method and device

    CN111583931A

  • Account opening method based on intelligent interactive questions and answers and related system

    CN112818104A