Data processing method and electronic device
By collecting and processing audio information and annotations in the projection mode, the system distinguishes between spoken and unspoken content, adjusts display parameters, solves the problem of speaker distraction, and improves the coherence and efficiency of the presentation.
Patent Information
- Application Number
- CN202210952283.4
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- Filing Date
- 2022-08-09
- Publication Date
- 2026-02-27
- Estimated Expiration
- 2042-08-09
AI Technical Summary
In screen mirroring mode, speakers need to frequently switch their attention between notes and presentation content while processing documents, leading to distraction and difficulty in locating content that has already been presented.
By collecting audio information from the currently displayed page, processing notes to distinguish between content that has been spoken and content that has not been spoken, and adjusting the display parameters of the notes, such as color and position, to ensure that the speaker can quickly locate the content that has not been spoken.
It effectively distinguishes between spoken and unspoken notes, helping speakers focus their attention and improve the coherence and efficiency of their presentations.
Smart Images

Figure CN115344224B_ABST
Abstract
Description
TECHNICAL FIELD
[0001] The present application relates to the technical field of data processing, in particular to a data processing method and an electronic device. BACKGROUND
[0002] In the screen projection mode, if there is a note information in the presentation file when the presenter is presenting the file, the attention of the presenter needs to be switched back and forth between the note information and the presentation content, thereby causing the attention of the presenter to be distracted, which is not conducive to the presenter to present the presentation content coherently or to insert new presentation content on the way, and the frequent visual switching makes it difficult for the presenter to locate the position where the note has been presented. SUMMARY
[0003] Therefore, the embodiments of the present application aim to provide a data processing method and an electronic device.
[0004] To achieve the above-mentioned purpose, the technical solution of the present application is as follows:
[0005] According to an aspect of the present application, a data processing method is provided, which comprises:
[0006] collecting voice information for a current presentation page;
[0007] obtaining note information in the current presentation page;
[0008] processing the note information based on the voice information to obtain first note content and second note content in the note information; wherein the first note content represents content that has been presented before the current time; and the second note content represents content that has not been presented before the current time;
[0009] adjusting a target parameter of the first note content or the second note content, so that the display content of the first note content and / or the second note content changes.
[0010] In the above-mentioned solution, adjusting the target parameter of the first note content or the second note content comprises one of the following:
[0011] adjusting a display parameter of the first note content or the second note content, so that the display mode of the first note content is different from that of the second note content;
[0012] adjusting a position parameter of the first note content or the second note content, so that the display order of the first note content is different from that of the second note content.
[0013] In the above-mentioned solution, adjusting the position parameter of the first note content comprises:
[0014] generate semantic content corresponding to the voice information;
[0015] match the semantic content with a plurality of sub-note information in the note information;
[0016] determine, according to a matching result, first sub-note information in the note information that matches the semantic content successfully; the first note content includes the first sub-note information;
[0017] adjust a note order of the first sub-note information in the note information, so that the first sub-note information is displayed before each sub-note information corresponding to the second note content.
[0018] The method further includes:
[0019] compare the semantic content with the first sub-note information;
[0020] determine, according to a comparison result, different content existing between the semantic content and the first sub-note information;
[0021] generate second sub-note information for the different content;
[0022] save the second sub-note information.
[0023] The method further includes:
[0024] output an inquiry request for adding the second sub-note information to the note information;
[0025] determine that response information received for the inquiry request is first response information;
[0026] add the second sub-note information to the note information;
[0027] The first response information indicates that the second sub-note information is agreed to be added to the note information.
[0028] The method further includes:
[0029] determine that response information received for the inquiry request is second response information;
[0030] delete the second sub-note information;
[0031] The second response information indicates that the second sub-note information is not agreed to be added to the note information.
[0032] The method further includes:
[0033] generate note display content corresponding to the second sub-note information;
[0034] outputting an inquiry request for adding the note presentation content to a presentation page corresponding to the current file;
[0035] determining that the response information received for the inquiry request is third response information;
[0036] adding the note presentation content to the presentation page corresponding to the current file;
[0037] wherein the third response information represents an agreement to add the note presentation content to the presentation page corresponding to the current file.
[0038] In the above solution, the obtaining of the note information in the current presentation page comprises:
[0039] determining that the note information does not exist in the current presentation page;
[0040] generating note information for the current presentation page based on the voice information.
[0041] In the above solution, the method further comprises:
[0042] receiving a note instruction for specific content in the current presentation page;
[0043] creating a note layer for the specific content in a first area of the specific content based on the note instruction;
[0044] displaying note information for the specific content in the note layer;
[0045] when the current presentation page is in a screen projection mode, controlling the note information of the specific content to be output only on a main display screen where the current presentation page is located.
[0046] According to another aspect of the present application, an electronic device is provided, comprising:
[0047] a collection unit configured to collect voice information for a current presentation page;
[0048] an obtaining unit configured to obtain note information in the current presentation page;
[0049] a processing unit configured to process the note information based on the voice information to obtain first note content and second note content in the note information; wherein the first note content represents content that has been narrated before the current time; and the second note content represents content that has not been narrated before the current time;
[0050] The adjustment unit is configured to adjust a target parameter of the first remark content or the second remark content, so that the display content of the first remark content and / or the second remark content is changed.
[0051] The data processing method and the electronic device provided in the present application can collect voice information for a current display page, obtain remark information of the current display page, process the remark information according to the voice information to obtain first remark content and second remark content in the remark information, and adjust a target parameter of the first remark content or the second remark content, so that the display content of the first remark content and / or the second remark content is changed. In this way, the content that has been narrated and the content that has not been narrated in the remark information can be distinguished, so that the speaker can quickly locate the content that has not been narrated. BRIEF DESCRIPTION OF DRAWINGS
[0052] Figure 1 The flowchart shows the implementation of the processing method in the present application;
[0053] Figure 2 The structural composition of the electronic device in the present application is shown in the flowchart Figure One ;
[0054] Figure 3 The structural composition of the electronic device in the present application is shown in the flowchart Figure Two . DETAILED DESCRIPTION
[0055] In order to make the purpose, technical solutions and advantages of the present application clearer, the technical solutions in the embodiments of the present application will be described clearly and completely in conjunction with the drawings in the embodiments of the present application. Obviously, the described embodiments are only some of the embodiments of the present application, but not all the embodiments. Based on the embodiments in the present application, all other embodiments obtained by those skilled in the art without creative labor fall within the scope of protection of the present application. In the case of no conflict, the embodiments in the present application and the features in the embodiments can be combined with each other at will. The steps shown in the flowchart of the drawings can be executed in a computer system such as a group of computer executable instructions. Moreover, although the logical order is shown in the flowchart, in some cases, the steps shown or described can be executed in an order different from that here.
[0056] As described above, since the attention of the speaker needs to be switched back and forth between the note information and the display when the presentation file (such as PPT) is projected, the attention is easily distracted and it is difficult to locate the note content that has been presented. The scheme of the present application collects the voice information and the note information for the current display page, processes the note information according to the voice information to obtain the first note content that has been presented and the second note content that has not been presented, and adjusts the target parameters of the first note content or the second note content to make the display content of the first note content and / or the second note content change. In this way, the presented content and the non-presented content in the note information can be distinguished, so that the speaker can quickly locate the non-presented content.
[0057] The technical scheme of the present application will be further described in detail below in combination with the drawings of the specification and specific embodiments.
[0058] Figure 1 The flowchart for implementing the data processing method in the present application is shown in the figure. The method can be applied to an electronic device with a display screen, including but not limited to a notebook computer, a desktop computer, a tablet computer, etc. The target file can be run in the electronic device, including but not limited to a presentation (PPT, PowerPoint), a WORD document generated by a word processing software (WORD, Microsoft office Word), etc. As shown in the figure, the method comprises the following steps. Figure 1
[0059] Step 101, collecting voice information for the current display page;
[0060] In the present application, the electronic device can be connected with a target device, and through the connection, the target file on the electronic device can be projected onto the secondary screen of the target device for display. When the target file is currently in the projection mode or the presentation mode, the electronic device can collect the voice information for the current display page of the target file through a voice collector.
[0061] Here, the voice collector can be a microphone provided by the electronic device, or a microphone independent of the electronic device.
[0062] Step 102, obtaining note information in the current display page;
[0063] In the present application, the target file can have note information, including but not limited to note information in a note area of the target file, note information in a display area. When the target file is currently in a screen projection mode or a presentation mode, the note information in the target file is only output on the main screen of the electronic device and will not be output on the secondary screen of the target device. In this way, the presenter can effectively present through the note information displayed on the display screen of the electronic device without affecting the audience or discussants.
[0064] In the present application, when the electronic device obtains the note information in the current display page, the note information can be obtained by extracting the note information in the note area or the display area of the current display page. If the note information is extracted in the note area or the display area of the current display page, the extracted note information can be directly processed subsequently; if the note information is not extracted in the note area or the display area of the current display page, it is determined that there is no note information in the current display page, and the electronic device can also generate note information for the current display page based on the voice information.
[0065] In one implementation, the electronic device can generate semantic content of the voice information and display the semantic content as the note information in the note area of the current display page.
[0066] In another implementation, the electronic device can extract key content from the semantic content after generating the semantic content of the voice information, and display the key content as the note information in the note area of the current display page. In this way, only the summary content of the voice information can be displayed in the note area, reducing the display and storage of invalid information.
[0067] In the present application, the electronic device can also detect the opening state of the microphone on the electronic device, and determine that the target file is currently in a screen projection mode or a presentation mode when it is detected that the microphone on the electronic device is in an open state.
[0068] In step 103, the note information is processed based on the voice information to obtain first note content and second note content in the note information; wherein the first note content represents content that has been presented before the current time; and the second note content represents content that has not been presented before the current time.
[0069] In the present application, after obtaining the note information of the current display page, the electronic device can classify the note information based on the voice information to divide the note information into first note content and second note content. The first note content represents content that has been presented before the current time; and the second note content represents content that has not been presented before the current time.
[0070] In an implementation, the electronic device can parse the voice information, generate semantic information of the voice information, match the semantic information with the note information, and according to a matching result, take note content that is successfully matched as the first note content, and take note content in the note information other than the first note content as the second note content.
[0071] At step 104, a target parameter of the first note content or the second note content is adjusted, so that display content of the first note content and / or the second note content is changed.
[0072] In the present application, the electronic device can adjust a display parameter of the first note content or the second note content, so that the first note content and the second note content are displayed in different manners.
[0073] In an implementation, the electronic device can darken (for example, to gray) a text color of the first note content, and the text color of the first note content remains unchanged. The first note content and the second note content are distinguished by changing the text color of the first note content.
[0074] In another implementation, the electronic device can highlight a text color of the second note content, and the text color of the first note content remains unchanged. The first note content and the second note content are distinguished by highlighting the second note content.
[0075] Here, the electronic device can help the speaker focus on the note content that has not been spoken by darkening the note content that has been spoken.
[0076] In the present application, the user may change a speaking order of note information in a current file when improvising during a speech, and therefore, the electronic device can also adjust a position parameter of the first note content or the second note content, so that the first note content and the second note content are displayed in different orders.
[0077] In an implementation, the electronic device can determine a current position order parameter of the first note content and the second note content, and by adjusting the current position order parameter of the first note content and the second note content, the position of the first note content is adjusted to be in front of the second note content.
[0078] For example, the remark information includes 10 sub-remark information, wherein, 1, 3, 5, 7, 9 sub-remark information belongs to the first remark content, 2, 4, 6, 8, 10 sub-remark information belongs to the second remark content, the electronic device can adjust the order of 1, 3, 5, 7, 9 sub-remark information in the 10 sub-remark information, that is, the original 1, 3, 5, 7, 9 sub-remark information is adjusted to 1, 2, 3, 4, 5 sub-remark information, and the original 2, 4, 6, 8, 10 sub-remark information is adjusted to 6, 7, 8, 9, 10 sub-remark information. In this way, the speech needs of different users can be met, and the coherence of the speech can be ensured.
[0079] In the present application, when adjusting the position parameter of the first remark content, the electronic device can also generate the semantic content corresponding to the voice information; match the semantic content with each of the plurality of sub-remark information in the remark information one by one; determine the first sub-remark information in the remark information that matches the semantic content successfully according to the matching result; here, the first remark content includes the first sub-remark information; adjust the remark order of the first sub-remark information in the remark information, so that the first sub-remark information is displayed before each sub-remark information corresponding to the second remark content.
[0080] For example, the remark information includes 10 sub-remark information, wherein, the semantic content corresponding to the current voice information matches the 5th sub-remark information successfully, and the remaining 1-4, 6-10 sub-remark information belongs to the second remark content, the electronic device can adjust the position of the 5th sub-remark information to the position of the first sub-remark information, and adjust the positions of the original 1-4, 6-10 sub-remark information to 2, 3, 4, 5, 6, 7, 8, 9, 10 sub-remark information, so as to ensure the coherence of the remark information.
[0081] In the present application, after adjusting the remark order of the remark information, the electronic device can also adjust the text color of the first remark content, for example, the text color of the first remark content is adjusted to dark gray, and the text color of the second remark content which has not been narrated remains unchanged, in this way, the coherence of the remark information can be ensured, and the speaker can quickly locate the content which has not been narrated, and the speech quality and efficiency of the speaker can be improved.
[0082] In the present application, since the speaker may produce impromptu speech content during the speech, this content may not exist in the remark information, therefore, the electronic device can also compare the semantic content with the first sub-remark information to obtain a comparison result; if there is different content between the semantic content and the first sub-remark information according to the comparison result, generate a second sub-remark information for the different content; and then save the second sub-remark information.
[0083] In the present application, the electronic device can further output an inquiry request for adding the second sub-note information to the note information in the case of the end of the speech; receive response information for the inquiry request; if it is determined that the received response information for the inquiry request is first response information; the electronic device can add the second sub-note information to the note information. Wherein, the first response information represents the consent to add the second sub-note information to the note information.
[0084] Here, the electronic device can determine that the speech of the electronic device is ended when receiving a closing instruction for the target file (i.e. the current speech file).
[0085] Alternatively, the electronic device can determine that the speech of the electronic device is ended when detecting that the text color of the note information has been adjusted (such as being adjusted to gray, representing the content that has been spoken).
[0086] Alternatively, the electronic device can determine that the speech of the electronic device is ended when detecting that there is a keyword (such as end of speech, thank you, END) in the current display of the target file.
[0087] In the present application, the electronic device can further delete the second sub-note information when it is determined that the received response information for the inquiry request is second response information; wherein, the second response information represents the disagreement to add the second sub-note information to the note information.
[0088] In this way, the scheme of the present application can perfect the note information by adding the improvisational speech content of the speaker as new note content to the note information, and facilitate the use next time.
[0089] In the present application, the electronic device can further generate note display content corresponding to the second sub-note information; then output an inquiry request for adding the note display content to the display page corresponding to the current file; receive response information for the inquiry request; and when it is determined that the received response information for the inquiry request is third response information, add the note display content to the display page corresponding to the current file.
[0090] Wherein, the third response information represents the consent to add the note display content to the display page corresponding to the current file.
[0091] In the present application, when it is determined that the received response information for the inquiry request is fourth response information, the note display content is deleted.
[0092] Wherein, the fourth response information represents the disagreement to add the note display content to the display page corresponding to the current file.
[0093] In this way, the user's impromptu speech can be added to the display page corresponding to the current file to improve the content of the current file and make it convenient for the speaker to use in the next speech.
[0094] In this application, the electronic device can also receive a note instruction for specific content on the currently displayed page; create a note layer for the specific content in a first area of the specific content based on the note instruction; and then display the note information for the specific content in the note layer.
[0095] Here, the electronic device can also control the annotation information of the specific content to be output only on the main display screen where the current display page is located when the current display page is in screen mirroring mode.
[0096] In other words, only the speaker can see the annotation layer and the annotation information within it, while the audience viewing the projection screen cannot see the annotation layer and the annotation information.
[0097] For example, for certain charts or formulas in the current presentation document, users can choose to create a note layer next to the chart or formula, and then add the corresponding note information in the note layer.
[0098] Here, the note layer will appear as an additional layer around the target chart or formula. This note layer is only displayed in the speaker view (i.e., the main display screen) and will not appear in the audience view (i.e., the secondary display screen connected to the main display screen).
[0099] Of course, this specific content can be the target text content, target image content, target code content, etc. selected by the user. It can be anything selected on the content display page of the target file in the current presentation.
[0100] In this application, the electronic device collects voice information about the currently displayed page; then obtains annotation information for the currently displayed page; processes the annotation information based on the voice information to obtain a first annotation and a second annotation; and then adjusts the target parameters of the first or second annotation to change the displayed content of the first and / or second annotation. In this way, it is possible to distinguish between previously mentioned and unmentioned content in the annotation information, allowing the speaker to quickly locate the unmentioned content.
[0101] Figure 2 This is a schematic diagram of the structural composition of the electronic device in this application. Figure One ,like Figure 2 As shown, the electronic device includes:
[0102] The acquisition unit 201 is used to acquire voice information for the currently displayed page;
[0103] The acquisition unit 202 is configured to acquire remark information in the current display page.
[0104] The processing unit 203 is configured to process the remark information based on the voice information, to obtain first remark content and second remark content in the remark information; the first remark content represents content that has been narrated before the current time; and the second remark content represents content that has not been narrated before the current time.
[0105] The adjustment unit 204 is configured to adjust a target parameter of the first remark content or the second remark content, so that display content of the first remark content and / or the second remark content changes.
[0106] In a preferred solution, the adjustment unit 204 can be specifically configured to adjust a display parameter of the first remark content or the second remark content, so that the first remark content and the second remark content are displayed in different manners; and adjust a position parameter of the first remark content or the second remark content, so that the first remark content and the second remark content are displayed in different orders.
[0107] In a preferred solution, the electronic device further includes:
[0108] The generation unit 205 is configured to generate semantic content corresponding to the voice information.
[0109] The matching unit 206 is configured to match the semantic content with multiple pieces of sub-remark information in the remark information.
[0110] The determination unit 207 is configured to determine, according to a matching result, first sub-remark information in the remark information that matches the semantic content successfully; and the first remark content includes the first sub-remark information.
[0111] The adjustment unit 204 is specifically configured to adjust a remark order of the first sub-remark information in the remark information, so that the first sub-remark information is displayed before each piece of sub-remark information corresponding to the second remark content.
[0112] In a preferred solution, the electronic device further includes a saving unit 208.
[0113] Specifically, the matching unit 206 is further configured to compare the semantic content with the first sub-remark information.
[0114] The determination unit 207 is further configured to determine, according to a comparison result, different content existing between the semantic content and the first sub-remark information.
[0115] The generating unit 205 is further configured to generate second sub-note information for the different content;
[0116] The saving unit 208 is configured to save the second sub-note information.
[0117] In a preferred implementation, the electronic device further includes an output unit 209 and an adding unit 210.
[0118] The output unit 209 is configured to output an inquiry request for adding the second sub-note information to the note information.
[0119] The determining unit 207 is further configured to determine that the response information received for the inquiry request is first response information.
[0120] The adding unit 210 is configured to add the second sub-note information to the note information.
[0121] The first response information indicates that the second sub-note information is agreed to be added to the note information.
[0122] In a preferred implementation, the electronic device further includes a deleting unit 211.
[0123] Specifically, the determining unit 207 is further configured to determine that the response information received for the inquiry request is second response information.
[0124] The deleting unit 211 is configured to delete the second sub-note information.
[0125] The second response information indicates that the second sub-note information is not agreed to be added to the note information.
[0126] In a preferred implementation, the generating unit 205 is further configured to generate note display content corresponding to the second sub-note information.
[0127] The output unit 209 is further configured to output an inquiry request for adding the note display content to a display page corresponding to a current file.
[0128] The determining unit 207 is further configured to determine that the response information received for the inquiry request is third response information.
[0129] The adding unit 210 is further configured to add the note display content to the display page corresponding to the current file.
[0130] The third response information indicates that the note display content is agreed to be added to the display page corresponding to the current file.
[0131] In a preferred implementation, the determining unit 207 is further configured to determine that the note information does not exist in a current display page.
[0132] The generating unit 205 is further configured to generate remark information for the current display page based on the voice information.
[0133] In a preferred implementation, the electronic device further comprises a receiving unit 212, a creating unit 213, a display unit 214 and a control unit 215.
[0134] The receiving unit 212 is configured to receive a remark instruction for specific content in the current display page.
[0135] The creating unit 213 is configured to create a remark layer for the specific content in a first area of the specific content based on the remark instruction.
[0136] The display unit 214 is configured to display remark information for the specific content in the remark layer.
[0137] The control unit 215 is configured to control the remark information of the specific content to be output only on the main display screen where the current display page is located when the current display page is in a projection mode.
[0138] It should be noted that the electronic device provided in the above embodiments is only used as an example for the division of the above program modules, and in actual applications, the above processing can be completed by different program modules according to needs, that is, the internal structure of the device is divided into different program modules to complete all or part of the above processing. In addition, the electronic device provided in the above embodiments and the processing method provided in the above embodiments belong to the same concept, and the specific implementation process is described in the method embodiments, which will not be repeated here.
[0139] The electronic device provided in the embodiments of the present application comprises a processor and a memory for storing a computer program capable of running on the processor,
[0140] The processor is configured to execute any method step of the above processing method when the computer program is running.
[0141] Figure 3 is a structural composition of the electronic device in the present application Figure Two The electronic device 300 can be a mobile phone, a computer, a digital broadcast terminal, an information transceiver device, a game console, a tablet device, a medical device, a fitness device, a personal digital assistant, etc. Figure 3The electronic device 300 shown includes at least one processor 301, a memory 302, at least one network interface 304, and a user interface 303. The various components in the electronic device 300 are coupled together by a bus system 305. As is known, the bus system 305 is used to facilitate communication among the components. The bus system 305 includes a data bus to facilitate the transfer of data between components, a control bus to facilitate the transfer of control information between components, and a status bus to facilitate the transfer of status information between components. However, for the sake of clarity, the various buses are shown as the bus system 305 in Figure 3
[0142] The user interface 303 can include a display, a keyboard, a mouse, a trackball, a click wheel, a keypad, a button, a touchpad, a touchscreen, etc.
[0143] It can be understood that the memory 302 can be a volatile memory or a non-volatile memory, and can also include both volatile and non-volatile memories. Among them, the non-volatile memory can be a read-only memory (ROM), a programmable read-only memory (PROM), an erasable programmable read-only memory (EPROM), an electrically erasable programmable read-only memory (EEPROM), a ferromagnetic random access memory (FRAM), a flash memory, a magnetic surface memory, an optical disc, or a compact disc read-only memory (CD-ROM); the magnetic surface memory can be a disk memory or a tape memory. The volatile memory can be a random access memory (RAM) used as an external cache. By way of example but not limitation, many forms of RAM can be used, such as static random access memory (SRAM), synchronous static random access memory (SSRAM), dynamic random access memory (DRAM), synchronous dynamic random access memory (SDRAM), double data rate synchronous dynamic random access memory (DDR SDRAM), enhanced synchronous dynamic random access memory (ESDRAM), synchronous link dynamic random access memory (SLDRAM), and direct rambus random access memory (DRRAM).The memory 302 described in the embodiments of the present application is intended to include, but is not limited to, these and any other suitable type of memory.
[0144] The memory 302 in the embodiments of the present application is used to store various types of data to support the operation of the electronic device 300. Examples of these data include: any computer programs used for operation on the electronic device 300, such as an operating system 3021 and an application program 3022; contact data; phonebook data; messages; pictures; audio; and the like. The operating system 3021 contains various system programs, such as a framework layer, a core library layer, a driver layer, and the like, for implementing various basic services and processing hardware-based tasks. The application program 3022 can contain various application programs, such as a media player (Media Player), a browser (Browser), and the like, for implementing various application services. The program implementing the method of the embodiments of the present application can be contained in the application program 3022.
[0145] The method disclosed in the embodiments of the present application can be applied in the processor 301 or implemented by the processor 301. The processor 301 can be an integrated circuit chip having a processing capability. In the implementation process, each step of the above method can be completed by the integrated logic circuit or the instruction in the form of software in the processor 301. The processor 301 described above can be a general-purpose processor, a digital signal processor (DSP), or other programmable logic device, discrete gate or transistor logic device, discrete hardware component, and the like. The processor 301 can implement or execute the methods, steps, and logic block diagrams disclosed in the embodiments of the present application. The general-purpose processor can be a microprocessor or any conventional processor, and the like. In combination with the steps of the method disclosed in the embodiments of the present application, the above-mentioned method can be directly embodied as a hardware coding processor to execute, or be executed by a combination of hardware and software modules in the coding processor. The software module can be located in the storage medium, and the storage medium is located in the memory 302. The processor 301 reads the information in the memory 302 and combines the hardware to complete the steps of the above-mentioned method.
[0146] In an exemplary embodiment, the electronic device 300 can be implemented by one or more Application Specific Integrated Circuits (ASICs), DSPs, Programmable Logic Devices (PLDs), Complex Programmable Logic Devices (CPLDs), Field-Programmable Gate Arrays (FPGAs), general-purpose processors, controllers, microcontrollers (MCUs), microprocessors (Microprocessors), or other electronic elements for executing the aforementioned methods.
[0147] In an exemplary embodiment, the embodiments of the present application further provide a computer readable storage medium, for example, the memory 302 including a computer program, which can be executed by the processor 301 of the electronic device 300 to complete the steps of the aforementioned methods. The computer readable storage medium can be a memory such as FRAM, ROM, PROM, EPROM, EEPROM, Flash Memory, magnetic surface memory, optical disc, or CD-ROM, etc.; or can be various devices including one or any combination of the above memories, such as mobile phones, computers, tablet devices, personal digital assistants, etc.
[0148] A computer readable storage medium having a computer program stored thereon, which, when executed by a processor, performs any of the steps of the above processing methods.
[0149] In several embodiments provided in the present application, it should be understood that the disclosed devices and methods can be implemented in other ways. The device embodiments described above are merely illustrative, for example, the division of the units is only a logical function division, and actual implementation can have another division manner, such as: multiple units or components can be combined, or can be integrated into another system, or some features can be ignored or not executed. In addition, the coupling or direct coupling or communication connection between the various components shown or discussed can be through some interfaces, indirect coupling or communication connection between devices or units, which can be electrical, mechanical or other forms.
[0150] The units described above as separate components can or can not be physically separated, and the components shown as units can or can not be physical units, that is, they can be located in one place or distributed on multiple network units; some or all of the units can be selected according to actual needs to achieve the purpose of the embodiments of the present application.
[0151] The methods disclosed in the several method embodiments provided by the present application can be combined arbitrarily without conflict to obtain new method embodiments.
[0152] The features disclosed in the several product embodiments provided by the present application can be combined arbitrarily without conflict to obtain new product embodiments.
[0153] The features disclosed in the several method or device embodiments provided by the present application can be combined arbitrarily without conflict to obtain new method embodiments or device embodiments.
[0154] The above description is merely a specific implementation of the present application, but the protection scope of the present application is not limited thereto. Any person skilled in the art can easily think of changes or replacements within the technical range disclosed by the present application, which should be covered within the protection scope of the present application. Therefore, the protection scope of the present application should be subject to the protection scope of the claims.
Claims
1. A data processing method, the method comprising: Collect voice information for the currently displayed page; Obtain the remarks information from the currently displayed page; The annotation information is processed based on the voice information to obtain a first annotation content and a second annotation content in the annotation information; wherein, the first annotation content represents the content that has been spoken before the current time; and the second annotation content represents the content that has not been spoken before the current time. Generate semantic content corresponding to the voice information; Based on the semantic content, adjust the display order of the first or second note content; The method further includes: The semantic content is compared with the first sub-note information; the first sub-note information is the note information that successfully matches the semantic content; the first note content includes the first sub-note information. Based on the comparison results, determine the differences between the semantic content and the first sub-note information; Generate a second sub-note information for each of the different contents; Save the second sub-note information.
2. The method according to claim 1, wherein, The method further includes: Match the semantic content with multiple sub-notes in the notes information; Based on the matching results, determine the first sub-note information in the note information that successfully matches the semantic content; Adjust the order of the first sub-note information in the note information so that the first sub-note information is displayed before each sub-note information corresponding to the second note content.
3. The method according to claim 1, wherein, The method further includes: Output a query request to add the second sub-note information to the note information; It is determined that the received response information in response to the inquiry request is the first response information; Add the second sub-note information to the note information; The first response information indicates agreement to add the second sub-note information to the note information.
4. The method according to claim 3, wherein, The method further includes: It is determined that the response information received in response to the inquiry request is the second response information; Delete the second sub-note information; The second response information indicates disagreement with adding the second sub-note information to the note information.
5. The method according to claim 1, wherein, The method further includes: Generate the comment display content corresponding to the second sub-comment information; Output a query request to add the aforementioned notes to the display page corresponding to the current file; It is determined that the response information received in response to the inquiry request is a third response information; Add the aforementioned notes to the display page corresponding to the current file; The third response information indicates agreement to add the remarks to the display page corresponding to the current file.
6. The method according to claim 1, wherein, The step of obtaining the remarks information in the currently displayed page includes: It has been confirmed that the aforementioned remarks do not exist on the currently displayed page; Based on the voice information, generate notes for the currently displayed page.
7. The method according to claim 1, wherein, The method further includes: Receive annotation instructions for specific content on the currently displayed page; Based on the annotation instruction, an annotation layer is created for the specific content in the first area of the specific content; The notes layer displays notes for the specific content. When the current display page is in screen mirroring mode, the annotation information of the specific content is controlled to be output only on the main display screen where the current display page is located.
8. An electronic device, comprising: The acquisition unit is used to collect voice information for the currently displayed page; The acquisition unit is used to acquire the remarks information in the currently displayed page; A processing unit is configured to process the annotation information based on the voice information to obtain a first annotation content and a second annotation content in the annotation information; wherein, the first annotation content represents content that has been spoken before the current time; and the second annotation content represents content that has not been spoken before the current time. A generation unit is used to generate semantic content corresponding to the speech information; An adjustment unit is used to adjust the display order of the first note content or the second note content based on the semantic content; A matching unit is configured to compare the semantic content with first sub-note information; the first sub-note information is note information that successfully matches the semantic content; the first note content includes the first sub-note information. A determining unit is configured to determine, based on the comparison result, the differences between the semantic content and the first sub-note information; The generation unit is also used to generate second sub-note information for the different content; A storage unit is used to store the second sub-note information.
Citation Information
Patent Citations
Speech content prompting method and system
CN112233669A
Providing prompts in real-time in speech recognition results
CN113763943A