Method, apparatus and video conferencing system for data display in video conferencing

By receiving and processing the reposted speech content in the video conferencing system and using historical video conferencing data to determine the relevant target speech data, the problem of inaccurate reposting in the video conferencing is solved, and the accuracy and comprehensiveness of the reposted content is improved.

CN117873628BActive Publication Date: 2025-05-27CHINA MOBILE INTERNET CO LTD +1
View PDF 1 Cites 0 Cited by

Patent Information

Application Number
CN202410051775.5
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2024-01-12
Publication Date
2025-05-27
Estimated Expiration
2044-01-12

AI Technical Summary

Technical Problem

In video conferences, because people speak quickly, participants’ records of the meeting content are not comprehensive and inaccurate, resulting in the content of the meeting being transferred to other personnel being incomplete and inaccurate.

Method used

By receiving the reposted speech content sent by the target video conferencing terminal device, the target speech data associated with the reposted speech content is determined based on the stored historical video conferencing data, and the target speech data is sent to the target video conferencing terminal device to display the target speech data by the target video conferencing terminal device.

Benefits of technology

It improves the accuracy and comprehensiveness of the content of user reposting, so that the target participants can repost the content of the meeting more accurately.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN117873628B_ABST
    Figure CN117873628B_ABST
Patent Text Reader

Abstract

The present application provides a method, apparatus, and video conferencing system for data display in a video conference, which relates to the technical field of video conferencing. Among them, the method applied to a video conferencing server includes: receiving the paraphrased speech content of a target participant in the current video conference sent by a target video conferencing terminal device; wherein, the target video conferencing terminal device is the video conferencing terminal device held by the target participant; the paraphrased speech content is the speech content of the target participant after the paraphrasing operation is triggered; determining target speech data associated with the paraphrased speech content based on the stored historical video conference data; the historical video conference data includes the speech data of each historical participant in the historical video conference; and sending the target speech data to the target video conferencing terminal device for the target video conferencing terminal device to display the target speech data. This solution can improve the accuracy and comprehensiveness of the content paraphrased by users.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] The present application relates to the technical field of video conferencing, and particularly to a method, apparatus, and video conferencing system for data display in a video conference. Background Art

[0002] Currently, video conferencing is being used more and more widely. More and more people hold meetings through video conferencing and record the content of the video conference to relay the meeting content to other people. However, since people speak relatively fast, participants may not record the meeting content comprehensively or accurately, resulting in incomplete and inaccurate meeting content relayed to other people. Summary of the Invention

[0003] To solve the above problems, the present application provides a method, apparatus, and video conferencing system for data display in a video conference.

[0004] According to a first aspect of the present application, there is provided a method for data display in a video conference, which is applied to a video conference server and includes:

[0005] Receiving the relayed speech content of a target participant in a current video conference sent by a target video conference terminal device; wherein, the target video conference terminal device is the video conference terminal device held by the target participant; the relayed speech content is the speech content of the target participant after the relaying operation is triggered;

[0006] Based on the stored historical video conference data, determining target speech data associated with the relayed speech content; the historical video conference data includes the speech data of each historical participant in the historical video conference;

[0007] Sending the target speech data to the target video conference terminal device for the target video conference terminal device to display the target speech data.

[0008] According to a second aspect of the present application, there is provided a method for data display in a video conference, which is applied to a video conference terminal device and includes:

[0009] Obtaining the relayed speech content of a target participant in a current video conference who holds the video conference terminal device; wherein, the relayed speech content is the speech content of the target participant after the relaying operation is triggered;

[0010] Sending the relayed speech content to a video conference server; wherein, the video conference server is used to determine target speech data associated with the relayed speech content based on the stored historical video conference data;

[0011] Receive the target speech data sent by the video server;

[0012] Display the target speech data in a target area on the display interface.

[0013] According to a third aspect of the present application, there is provided a device for data display in a video conference, which is applied to a video conference server and includes:

[0014] A first receiving module, configured to receive the paraphrased speech content of a target participant in the current video conference sent by a target video conference terminal device; wherein, the target video conference terminal device is the video conference terminal device held by the target participant; the paraphrased speech content is the speech content of the target participant after the paraphrasing operation is triggered;

[0015] A first determining module, configured to determine target speech data associated with the paraphrased speech content based on the stored historical video conference data; the historical video conference data includes the speech data of each historical participant in the historical video conference;

[0016] A first sending module, configured to send the target speech data to the target video conference terminal device for the target video conference terminal device to display the target speech data.

[0017] According to a fourth aspect of the present application, there is provided a device for data display in a video conference, which is applied to a video conference terminal device and includes:

[0018] A first obtaining module, configured to obtain the paraphrased speech content of a target participant holding the video conference terminal device in the current video conference; wherein, the paraphrased speech content is the speech content of the target participant after the paraphrasing operation is triggered;

[0019] A second sending module, configured to send the paraphrased speech content to a video conference server; wherein, the video conference server is configured to determine target speech data associated with the paraphrased speech content based on the stored historical video conference data;

[0020] A second receiving module, configured to receive the target speech data sent by the video server;

[0021] A display module, configured to display the target speech data in a target area on the display interface.

[0022] According to a fifth aspect of the present application, there is provided a video conference system, including:

[0023] A video conference server, configured to execute the method described in the first aspect above;

[0024] A video conferencing terminal device is configured to execute the method described in the second aspect above.

[0025] According to a sixth aspect of the present application, there is provided an electronic device, including: a processor; a memory for storing executable instructions of the processor; wherein, the processor is configured to execute the instructions to implement the method described in the first aspect above, and / or implement the method described in the second aspect above.

[0026] According to a seventh aspect of the present application, there is provided a computer-readable storage medium, when the instructions in the computer-readable storage medium are executed by a processor of an electronic device, enabling the electronic device to execute the method described in the first aspect above, and / or execute the method described in the second aspect above.

[0027] According to the technical solution of the present application, in a video conference, if a target participant needs to relay the conference content to others, the video conference server can determine target speech data associated with the relayed speech content of the target participant based on the stored historical video conference data, and send the target speech data to the target video conferencing terminal device for the target video conferencing terminal device to display the target speech data. In this way, the target participant can relay the conference content based on the target speech data, thereby greatly improving the accuracy and comprehensiveness of the content relayed by the user.

[0028] Additional aspects and advantages of the present application will be given in part in the following description, become apparent in part from the following description, or be understood through the practice of the present application. BRIEF DESCRIPTION OF THE DRAWINGS

[0029] The above and / or additional aspects and advantages of the present application will become apparent and be readily understood from the following description of the embodiments in conjunction with the drawings, wherein:

[0030] Figure 1 is a flowchart of a method for data display in a video conference provided by an embodiment of the present application;

[0031] Figure 2 is a flowchart of another method for data display in a video conference provided by an embodiment of the present application;

[0032] Figure 3 is a flowchart of yet another method for data display in a video conference provided by an embodiment of the present application;

[0033] Figure 4 is a flowchart of yet another method for data display in a video conference provided by an embodiment of the present application;

[0034] Figure 5Flowchart of another method for data display in a video conference provided by an embodiment of the present application;

[0035] Figure 6 Flowchart of another method for data display in a video conference provided by an embodiment of the present application;

[0036] Figure 7 Block diagram of a device for data display in a video conference provided by an embodiment of the present application;

[0037] Figure 8 Block diagram of another device for data display in a video conference provided by an embodiment of the present application;

[0038] Figure 9 Block diagram of a video conference system provided by an embodiment of the present application;

[0039] Figure 10 Block diagram of an electronic device provided by an embodiment of the present application. Detailed implementation manners

[0040] The embodiments of the present application are described in detail below. The examples of the embodiments are shown in the accompanying drawings, where the same or similar reference numerals denote the same or similar elements or elements having the same or similar functions from beginning to end. The embodiments described below with reference to the accompanying drawings are exemplary and are intended to explain the present application, but should not be construed as limiting the present application.

[0041] In the technical solution of the present application, the acquisition, storage, and application of the user's personal information involved all comply with the provisions of relevant laws and regulations and do not violate public order and good customs. The user's personal information involved is acquired, stored, and applied with the consent of the user.

[0042] It should be noted that currently, video conferences are applied more and more widely. More and more people hold meetings through video conferences and record the content of the video conferences to relay the meeting content to other people. However, since people speak relatively fast, the participants may not record the meeting content comprehensively and accurately, resulting in the meeting content relayed to other people being incomplete and inaccurate.

[0043] To solve the above problems, the present application proposes a method, a device, and a video conference system for data display in a video conference.

[0044] Figure 1The flowchart of a method for data display in a video conference provided by an embodiment of the present application. The method for data display in a video conference provided by an embodiment of the present application can be applied to a video conference server. It should be noted that the method for data display in a video conference in an embodiment of the present application can also be used in the device for data display in a video conference in an embodiment of the present application, and the device can be configured in an electronic device. As Figure 1 shown, the method may include the following steps:

[0045] Step 101, receive the paraphrased speech content of a target participant in the current video conference sent by a target video conference terminal device; wherein, the target video conference terminal device is the video conference terminal device held by the target participant; the paraphrased speech content is the speech content of the target participant after the paraphrasing operation is triggered.

[0046] It can be understood that when conducting a video conference, multiple participants are required to participate, and each participant holds a video conference terminal device to participate in the video conference.

[0047] In some embodiments of the present application, the current video conference refers to the video conference currently in progress. The target participant refers to the participant who triggers the paraphrasing operation of the target video conference terminal device in the current video conference. The target video conference terminal device is held by the target participant and is used for the current video conference. The paraphrased speech content can be data in the form of voice or data in the form of text.

[0048] As an example, when a target participant needs to paraphrase the meeting content to other participants, the paraphrasing operation can be triggered, and after the paraphrasing operation is triggered, information related to the meeting content to be paraphrased is spoken. After the paraphrasing operation ends, the target video conference terminal device takes the speech content of the target participant after the paraphrasing operation is triggered as the paraphrased speech content and sends the paraphrased speech content to the video conference server.

[0049] Step 102, based on the stored historical video conference data, determine the target speech data associated with the paraphrased speech content; the historical video conference data includes the speech data of each historical participant in the historical video conference.

[0050] Among them, the target speech content refers to the speech content in the historical meeting associated with the paraphrased speech content, that is, the speech content that the target participant needs to paraphrase.

[0051] In some embodiments of the present application, the historical video conference data includes the speech data of each historical participant in the historical video conference. Among them, the speech data of each historical participant may include the speech content in voice form, the semantic recognition text data corresponding to the speech content in voice form, and may also include image data and video data during the speech process, etc. When determining the target speech data, if the paraphrased speech content is data in voice form, the paraphrased speech content can be semantically recognized first to obtain the corresponding semantic recognition text data; the semantic recognition text data corresponding to the paraphrased speech content is semantically matched with the semantic recognition text data in the historical video conference data, and the speech data of the historical participant corresponding to the semantic recognition text data with the highest matching value is determined as the target speech data.

[0052] As an implementation manner, the historical video conference data can be stored in the form of a blockchain. For example, the video conference server generates the head block of the blockchain for the historical video conference. Each video conference terminal device records the speech content of the participants and generates a block according to the speech content, and uploads the corresponding block to the video conference server; after the video conference server processes these data, a new block is generated and incorporated into the chain.

[0053] Step 103, send the target speech data to the target video conference terminal device for the target video conference terminal device to display the target speech data.

[0054] That is to say, when the target participant needs to paraphrase the content, the target speech data in the historical video conference can be matched based on the paraphrased speech content of the target participant. That is, the target speech data is the speech data to be paraphrased matched based on the paraphrased speech content. By sending the target speech data to the target video conference terminal device, the target video conference terminal device can display the target speech data, so that the target participant can paraphrase the conference content based on the displayed target speech data, thereby improving the accuracy and comprehensiveness of the conference content paraphrase.

[0055] According to the method for data display in a video conference according to the embodiments of the present application, the video conference server receives the paraphrased speech content of the target participant in the current video conference sent by the target video conference terminal device, determines the target speech data associated with the paraphrased speech content based on the stored historical video conference data, and sends the target speech data to the target video conference terminal device for the target video conference terminal device to display the target speech data. In this way, the target participant can paraphrase the conference content based on the target speech data, thereby greatly improving the accuracy and comprehensiveness of the content paraphrased by the user.

[0056] Next, the implementation process of determining the target speech data associated with the paraphrased speech content will be introduced in detail.

[0057] Figure 2 It is a flowchart of another method for data display in a video conference provided by an embodiment of the present application. As Figure 2 shown, based on the above embodiment, Figure 1 the implementation process of step 102 in

[0058] Step 201, determine the target participant corresponding to the paraphrased speech content.

[0059] Step 202, based on the stored historical video conference data, determine the target historical video conference participated by the target participant.

[0060] It should be noted that the historical video conference data may include the speech data of each participant in each historical video conference, where the speech data includes the information of the participant and may also include information such as the speech time.

[0061] It can be understood that the prerequisite for a target participant to paraphrase the content of a certain historical video conference is that the target participant also participated in that historical video conference. Therefore, the target speech data can be determined from the target historical video conference participated by the target participant.

[0062] Step 203, based on the stored historical video conference data, determine the matching value between the paraphrased speech content and the speech data of each historical participant in the target historical video conference.

[0063] In some embodiments of the present application, the historical video conference data includes the speech data of each historical participant in each historical video conference. Among them, the speech data of each historical participant may include the speech content in text form, the speech content in voice form, the semantic recognition text data corresponding to the speech content in voice form, and may also include the speech content in video or image form, and may also include the information of the historical participant, the speech time, etc.

[0064] As an example, when determining the matching value between the paraphrased speech content and the speech data of each historical participant in the target historical video conference, if the paraphrased speech content is in text form, the matching value between the paraphrased speech content and the text form of the speech content or the semantic recognition text data in the speech data of each historical participant can be determined through semantic matching. Among them, the specific implementation method of semantic matching can be realized through the semantic matching model in related technologies, or through the semantic matching model constructed by relevant technical personnel in this field based on the training samples composed of the paraphrased speech content and the speech data of historical participants, or through other relevant algorithms to calculate the matching value.

[0065] As another example, if the paraphrased speech content is in speech form, semantic recognition can be performed on the paraphrased speech content to obtain the first semantic recognition text data corresponding to the paraphrased speech content; the matching value between the first semantic recognition text data and the text form of the speech content or the semantic recognition text data in the speech data of each historical participant can be determined through semantic matching.

[0066] Step 204: Determine the target speech data associated with the paraphrased speech content according to the matching value.

[0067] As an implementation method, if the target speech data only includes a set of data, the speech data of historical participants can be sorted in descending order according to the matching value, and the speech data corresponding to the first matching value in the sorting can be determined as the target speech data.

[0068] As another implementation method, if the target speech data can include multiple sets of data, the matching value can be compared with a preset matching degree threshold, and the speech data of historical participants with a matching value greater than the matching degree threshold can be determined as the target speech data.

[0069] In some other embodiments of the present application, in order to protect the privacy of historical participants, the speech data of each historical participant in the historical video conference data may further include permission information. The implementation process of determining the target speech data associated with the paraphrased speech content according to the matching value may include: determining the permission information in the speech data of each historical participant in the target historical video conference; according to the permission information, determining the first target speech data from the speech data of each historical participant in the target historical video conference, and according to the matching value, determining the target speech data associated with the paraphrased speech content from the first target speech data.

[0070] As an example, if the permission information of the speech data of historical participants can include the identification information of the participants allowed to view, the paraphrased speech content can include the identification information of the target participant. In this way, according to the permission information, the first target speech data allowed for the target participant to view can be determined from the speech data of historical participants.

[0071] According to the method for data display in a video conference according to an embodiment of the present application, the video conference server first determines the target historical video conference participated by the target participant based on the stored historical video conference data, then determines the matching value between the paraphrased speech content and the speech data of each historical participant in the target historical video conference, and determines the target speech data associated with the paraphrased speech content according to the matching value. This solution can match the paraphrased speech content with the speech data in the historical video conference data to obtain the target speech data associated with the paraphrased speech content, and then send the target speech data to the target video conference terminal device, so that the target participant can paraphrase the meeting content based on the target speech data, thereby greatly improving the accuracy and comprehensiveness of the content paraphrased by the user.

[0072] It should be noted that the historical video conference data is continuously updated, and the speech data of each participant in the current video conference also needs to be stored in the historical video conference data. Next, the storage process of the speech data of each participant in the current video conference will be introduced.

[0073] Figure 3 It is a flowchart of another method for data display in a video conference provided by an embodiment of the present application. As Figure 3 shown, based on the above embodiment, the method further includes the following steps:

[0074] Step 301, receive the speech content of each participant sent by multiple video conference terminal devices corresponding to the current video conference.

[0075] Among them, the multiple video conference terminal devices corresponding to the current video conference refer to the multiple video conference terminal devices participating in the current video conference. For example, when the current video conference ends, the multiple video conference terminal devices participating in the current video conference send the speech content of the corresponding participants to the video conference server. Among them, the speech content of the participants can include the speech content in voice form, can also include the speech content in text form, can also include the speech content in the form of images or videos, and can also include the speech time, participant information, etc.

[0076] In some embodiments of the present application, the participants can set permissions for their speech content, and the permission information of the speech content is sent to the video conferencing server together with the speech content. That is to say, step 301 may further include: receiving the speech content of each participant and the corresponding permission information of the speech content sent by a plurality of video conferencing terminal devices corresponding to the current video conference.

[0077] Step 302, perform semantic recognition on the speech content to obtain corresponding semantic recognition text data.

[0078] In some embodiments of the present application, semantic recognition can be performed on the speech content in the form of speech in the speech content to obtain corresponding semantic recognition text data.

[0079] Step 303, determine the speech content and the semantic recognition text data as the speech data of the corresponding participant.

[0080] If the data received from each video conferencing terminal device also includes the permission information corresponding to the speech content of each participant, then the speech content, the permission information corresponding to the speech content, and the semantic recognition text data are determined as the speech data of the corresponding participant.

[0081] Step 304, store the speech data of each participant into the historical video conference data.

[0082] That is to say, after the current video conference ends, the speech data of each participant in the current video conference will also be used as the speech data of the historical participants in the historical video conference and stored in the historical video conference data. Among them, the speech data of each historical participant in each historical video conference is also stored in the historical video conference data in the above manner.

[0083] According to the method for displaying data in a video conference according to an embodiment of the present application, the video conferencing server performs semantic recognition on the speech content of each participant sent by a plurality of video conferencing terminal devices corresponding to the current video conference to obtain corresponding semantic recognition text data, and determines the speech content and the semantic recognition text data as the speech data of the corresponding participant, and stores the speech data of each participant into the historical video conference data, so as to construct the historical video conference data, thereby enabling the associated target speech data to be retrieved for the paraphrased speech content and sending the target speech data to the target video conferencing terminal device to improve the accuracy and comprehensiveness of the target participant's paraphrased conference content.

[0084] In order to improve the accuracy of semantic recognition, another embodiment is proposed in the present application.

[0085] Figure 4The flowchart of another method for data display in a video conference provided by an embodiment of this application. As Figure 4 shown, based on the above embodiment, Figure 3 the implementation process of step 302 in

[0086] Step 401: Perform semantic recognition on the speech content to obtain the corresponding text data of the semantic recognition to be confirmed.

[0087] Step 402: Based on the text data of the semantic recognition to be confirmed, send a text data confirmation request to the first video conference terminal device held by the corresponding participant.

[0088] Among them, the text data confirmation request may include the speech content in voice form and also include the corresponding text data of the semantic recognition to be confirmed.

[0089] Step 403: Receive the confirmation result sent by the first video conference terminal device, and determine the corresponding text data of the semantic recognition according to the confirmation result.

[0090] That is to say, after semantic recognition, the semantic recognition result will be sent to the corresponding video conference terminal device for confirmation by the corresponding participant, and then according to the confirmation result, the text data of the semantic recognition corresponding to the speech data will be determined to ensure the accuracy of the semantic recognition result, so as to improve the correctness of the speech data in the historical video conference data, and further improve the accuracy of the participants' retelling of the meeting content.

[0091] According to the method for data display in a video conference of an embodiment of this application, by performing semantic recognition on the speech content to obtain the corresponding text data of the semantic recognition to be confirmed, based on the text data of the semantic recognition to be confirmed, sending a text data confirmation request to the first video conference terminal device held by the corresponding participant, and receiving the confirmation result sent by the first video conference terminal device, and determining the corresponding text data of the semantic recognition according to the confirmation result. This solution can confirm the text data of the semantic recognition to be confirmed by the corresponding participant to improve the accuracy of the text data of the semantic recognition, so as to improve the correctness of the speech data in the historical video conference data, and further improve the accuracy of the participants' retelling of the meeting content.

[0092] Next, the implementation process of the method for data display in a video conference applied to a video conference terminal device will be introduced.

[0093] Figure 5The flowchart of another method for data display in a video conference provided by an embodiment of the present application. The method for data display in a video conference according to the embodiment of the present application is applied to a video conference terminal device. It should be noted that the method for data display in a video conference according to the embodiment of the present application can be applied to the device for data display in a video conference provided by the embodiment of the present application, and this device can be configured in an electronic device. As Figure 5 shown, the method may include the following steps:

[0094] Step 501, obtain the paraphrased speech content of the target participant who holds the video conference terminal device in the current video conference; wherein, the paraphrased speech content is the speech content of the target participant after the paraphrasing operation is triggered.

[0095] That is to say, after the paraphrasing operation is triggered, continuously collect the speech content of the target participant, and use the collected speech content as the paraphrased speech content. Among them, the paraphrasing operation can be the triggering of an operation component in the display interface of the video conference terminal device. For example, there is a button for paraphrased speech displayed in the display interface. The target participant can touch the button and say information related to the content to be paraphrased, and stop touching the button after the paraphrased speech.

[0096] In some embodiments of the present application, the paraphrased speech content can be speech content in voice form or text form, which is not limited here.

[0097] As an example, during the speech of the target participant, if it is necessary to paraphrase the content of a historical video conference, the paraphrasing operation can be triggered, and the information of the content to be paraphrased can be expressed during the speech when the paraphrasing operation is triggered. The video conference terminal device will use the speech content collected after the paraphrasing operation is triggered as the paraphrased speech content.

[0098] Step 502, send the paraphrased speech content to the video conference server; wherein, the video conference server is used to determine the target speech data associated with the paraphrased speech content based on the stored historical video conference data.

[0099] Step 503, receive the target speech data sent by the video server.

[0100] Among them, the target speech data refers to the speech data in the historical video conference associated with the paraphrased speech content, that is, the reference data for the target participant to conduct a meeting paraphrase.

[0101] Step 504, display the target speech data in the target area in the display interface.

[0102] In some embodiments of the present application, the target speech data may include speech content in the form of speech, and may also include semantic recognition text data corresponding to the speech data in the form of speech. When displaying the target speech data in a target area on the display interface, the semantic recognition text data in the target speech data may be displayed in the target area, so that the target meeting participants can refer to the content displayed on the display interface when relaying the meeting content, thereby improving the accuracy of the relay.

[0103] In some other embodiments of the present application, the number of groups of target speech data may be multiple. Therefore, when displaying the target speech data in a target area on the display interface, the multiple groups of target speech data may be scrolled and played in the target area, so that the target meeting participants can view all the target speech data. In addition, the target speech data may further include information such as the corresponding meeting participant's information and speech time, and the display interface may further include a selection and confirmation operation for each group of target speech data. If the target meeting participant selects a certain group of target speech data, the other target speech data will no longer be displayed, and the display interface will only display the selected group of target speech data.

[0104] It should be noted that in addition to including speech content in text form and speech content in speech form, the target speech data may also include speech content in the form of images or videos. The target meeting participants can select and configure the speech content to be displayed. For example, they can configure to only display text content, or can also configure to display image content or video content.

[0105] In some other embodiments of the present application, when the target meeting participants relay the content of a historical meeting, in order to enable the relevant meeting participants to understand the relayed content in detail, there is a need to share the target speech data with other meeting participants in the current video conference. To meet this need, the video conference terminal device may also be equipped with a sharing function for the target speech data. As an implementation method, the present method may further include: in response to a sharing trigger operation for the target speech data, determining the meeting participants to be shared according to the sharing trigger operation; and sending the target speech data to the video conference terminal device held by the meeting participants to be shared.

[0106] As another implementation method, the target speech data may further include permission information, and the permission information may include the identification information of the meeting participants allowed to view the target speech data. Therefore, when sharing the target speech data with other meeting participants, it can only be shared with the meeting participants with permissions. That is to say, the implementation method of sending the target speech data to the video conference terminal device held by the meeting participants to be shared includes: determining the target sharing meeting participants with sharing permissions from the meeting participants to be shared according to the permission information; and sending the target speech data to the video conference terminal device held by the target sharing meeting participants.

[0107] According to the method for data display in a video conference according to an embodiment of the present application, the video conference terminal device obtains the paraphrased speech content of a target participant in the current video conference, and sends the paraphrased speech content to the video conference server, so that the video conference server determines target speech data associated with the paraphrased speech content based on historical video conference data, receives the target speech data sent by the video conference server, and displays the target speech data in a target area on the display interface. In this way, the target participant can paraphrase the meeting content based on the target speech data, thereby greatly improving the accuracy and comprehensiveness of the content paraphrased by the user.

[0108] Figure 6 It is a flowchart of another method for data display in a video conference provided by an embodiment of the present application. As Figure 6 shown, based on the above embodiment, the method may further include:

[0109] Step 601, sending the speech content of the target participant in the current video conference to the video conference server; after receiving the speech content, the video conference server performs semantic recognition on the speech content, and stores the obtained semantic recognition text data and the speech content as the speech data of the target participant in the historical video conference data.

[0110] As a possible implementation manner, the video conference terminal device may send the speech content of the target participant in the current video conference to the video conference server when the current video conference ends. As an example, during the conference, the video conference terminal device may store the speech content of the target participant, and generate a block using a private key signature, and after the current video conference ends, send the generated block to the video conference server.

[0111] As another possible implementation manner, the video conference terminal device may also send the corresponding speech content to the video conference server each time the target participant finishes speaking, or may send the speech content of the target participant during this period at preset time intervals.

[0112] In some embodiments of the present application, the display interface of the video conference terminal device is equipped with a permission configuration component, and the target participant can also set corresponding permissions for the speech content through the permission configuration component, such as setting the participants who can view the speech content through the permission configuration component. That is to say, the method may further include: in response to a trigger operation of the permission configuration component, determining the permission information corresponding to the speech content; and sending the permission information corresponding to the speech content and the speech content to the video conference server together.

[0113] In some other embodiments of the present application, after the video conferencing server performs semantic recognition on the received speech content, it is also necessary to send a text data confirmation request to the video conferencing terminal device. In this case, the method further includes:

[0114] Step 602, receiving the text data confirmation request sent by the video conferencing server.

[0115] Step 603, determining the semantic recognition text data to be confirmed according to the text data confirmation request, and displaying the semantic recognition text data to be confirmed in the corresponding area of the display interface for the target participant to perform a confirmation operation.

[0116] Among them, the text data confirmation request may include the speech content in voice form and the corresponding semantic recognition text data to be confirmed. After receiving the text data confirmation request, the video conferencing terminal device can determine the semantic recognition text data to be confirmed and the corresponding speech content, and display the semantic recognition text data to be confirmed in the corresponding area of the display interface for the target participant to perform a confirmation operation. In addition, the speech content can also be displayed in the form of an operation component on the display interface. When the target participant needs to play the speech content, the operation component can be triggered, and the video conferencing terminal device plays the speech content after receiving the trigger operation of the operation component.

[0117] As another implementation method, the video conferencing terminal device can also display the text data confirmation request in the form of a message, and when the target participant triggers the message, then display the semantic recognition text data to be confirmed in the corresponding area of the display interface for the target participant to perform a confirmation operation.

[0118] Step 604, in response to the confirmation operation of the target participant on the semantic recognition text data to be confirmed, obtaining a confirmation result and sending the confirmation result to the video conferencing server.

[0119] During the confirmation process of the target participant, the display interface of the video conferencing terminal device may include a content modification area and a confirmation submission button. After the confirmation submission button is triggered, the video conferencing terminal device obtains the confirmed semantic recognition text data according to the content in the modification area, and sends the confirmed semantic recognition text data as the confirmation result to the video conferencing server.

[0120] According to the method for data display in a video conference according to an embodiment of the present application, the video conference terminal device sends the speech content of the target participant in the current video conference to the video conference server, so that the video conference server performs semantic recognition on the speech content, and stores the obtained semantic recognition text data and the speech content as the speech data of the target participant in the historical video conference data, thereby realizing the storage of the speech content of each participant in the current video conference to ensure the accuracy when the participants relay the conference content. In addition, according to the text data confirmation request, the to-be-confirmed semantic recognition text data is displayed on the display interface, and after the target participant confirms the to-be-confirmed semantic recognition text data, the confirmation result is sent to the video conference server to ensure the accuracy of the data stored in the historical video conference data, thereby further improving the accuracy of the participants relaying the conference content.

[0121] To implement the above embodiment, the present application provides a device for data display in a video conference.

[0122] Figure 7 The following is a structural block diagram of a device for data display in a video conference provided by an embodiment of the present application. The device for data display in a video conference according to an embodiment of the present application is applied to a video conference server. As Figure 7 shown, the device includes:

[0123] A first receiving module 701, configured to receive the relay speech content of the target participant in the current video conference sent by the target video conference terminal device; wherein, the target video conference terminal device is the video conference terminal device held by the target participant; the relay speech content is the speech content of the target participant after the relay operation is triggered;

[0124] A first determining module 702, configured to determine the target speech data associated with the relay speech content based on the stored historical video conference data; the historical video conference data includes the speech data of each historical participant in the historical video conference;

[0125] A first sending module 703, configured to send the target speech data to the target video conference terminal device for the target video conference terminal device to display the target speech data.

[0126] In some embodiments of the present application, the first determining module 702 is specifically configured to:

[0127] Determine the target participant corresponding to the relay speech content;

[0128] Based on the historical video conference data, determine the target historical video conference participated by the target participant;

[0129] Based on historical video conference data, determine the matching value between the paraphrased speech content and the speech data of each historical participant in the target historical video conference;

[0130] According to the matching value, determine the target speech data associated with the paraphrased speech content.

[0131] As a possible implementation, the first determination module 702 is further configured to:

[0132] Determine the permission information in the speech data of each historical participant in the target historical video conference;

[0133] According to the permission information, determine the first target speech data from the speech data of each historical participant in the target historical video conference;

[0134] According to the matching value, determine the target speech data associated with the paraphrased speech content from the first target speech data.

[0135] As another possible implementation, the first determination module 702 is further configured to:

[0136] Compare the matching value with a preset matching degree threshold, and determine the speech data of the historical participant whose matching value is greater than the matching degree threshold as the target speech data.

[0137] In some embodiments of the present application, the apparatus further includes:

[0138] A third receiving module 704, configured to receive the speech content of each participant sent by a plurality of video conference terminal devices corresponding to the current video conference;

[0139] A second obtaining module 705, configured to perform semantic recognition on the speech content to obtain corresponding semantic recognition text data;

[0140] A second determination module 706, configured to determine the speech content and the semantic recognition text data as the speech data of the corresponding participant;

[0141] A storage module 707, configured to store the speech data of each participant in the historical video conference data.

[0142] As a possible implementation, the second obtaining module 705 is specifically configured to:

[0143] Perform semantic recognition on the speech content to obtain corresponding to-be-confirmed semantic recognition text data;

[0144] Based on the to-be-confirmed semantic recognition text data, send a text data confirmation request to the first video conference terminal device held by the corresponding participant;

[0145] Receive the confirmation result sent by the first video conferencing terminal device, and determine the corresponding semantic recognition text data according to the confirmation result.

[0146] As a possible implementation manner, the third receiving module 704 is further configured to:

[0147] Receive the speech content of each participant and the corresponding permission information of the speech content sent by multiple video conferencing terminal devices corresponding to the current video conference;

[0148] Wherein, the second determination module 706 is further configured to:

[0149] Determine the speech content, the corresponding permission information of the speech content, and the semantic recognition text data as the speech data of the corresponding participant.

[0150] According to the device for data display in a video conference according to an embodiment of the present application, the video conference server receives the reported speech content of the target participant in the current video conference sent by the target video conferencing terminal device, and determines the target speech data associated with the reported speech content based on the stored historical video conference data, and sends the target speech data to the target video conferencing terminal device for the target video conferencing terminal device to display the target speech data. In this way, the target participant can report the conference content based on the target speech data, thereby greatly improving the accuracy and comprehensiveness of the user's reported content.

[0151] Figure 8 It is a structural block diagram of another device for data display in a video conference provided by an embodiment of the present application. It should be noted that this device is applied to a video conferencing terminal device. As Figure 8 shown, the device includes:

[0152] The first acquisition module 801 is configured to acquire the reported speech content of the target participant in the current video conference who holds the video conferencing terminal device; wherein, the reported speech content is the speech content of the target participant after the reporting operation is triggered;

[0153] The second sending module 802 is configured to send the reported speech content to the video conference server; wherein, the video conference server is configured to determine the target speech data associated with the reported speech content based on the stored historical video conference data;

[0154] The second receiving module 803 is configured to receive the target speech data sent by the video server;

[0155] The display module 804 is configured to display the target speech data in a target area on the display interface.

[0156] In some embodiments of the present application, the display module 804 is specifically configured to:

[0157] Display the semantic recognition text data in the target speech data in the target area.

[0158] As an implementation, if the number of target speech data is multiple groups; the display module is further configured to:

[0159] Scroll and play and display multiple groups of target speech data in the target area.

[0160] In some embodiments of the present application, the device further includes:

[0161] A sharing module 805, configured to determine the participants to be shared according to the sharing trigger operation in response to the sharing trigger operation of the target speech data;

[0162] A third sending module 806, configured to send the target speech data to the video conferencing terminal device held by the participants to be shared.

[0163] In some embodiments of the present application, the target speech data includes permission information; the third sending module 806 is further configured to:

[0164] Determine the target sharing participants with sharing permissions from the participants to be shared according to the permission information;

[0165] Send the target speech data to the video conferencing terminal device held by the target sharing participants.

[0166] In some embodiments of the present application, the device further includes:

[0167] A fourth sending module 807, configured to send the speech content of the target participants in the current video conference to the video conference server; after receiving the speech content, the video conference server performs semantic recognition on the speech content, and stores the obtained semantic recognition text data and the speech content as the speech data of the target participants in the historical video conference data.

[0168] As a possible implementation, the device further includes a confirmation module 808, and the confirmation module 808 is specifically configured to:

[0169] Receive a text data confirmation request sent by the video conference server;

[0170] Determine the semantic recognition text data to be confirmed according to the text data confirmation request, and display the semantic recognition text data to be confirmed in the corresponding area of the display interface for the target participants to perform a confirmation operation;

[0171] In response to the confirmation operation of the target participants on the semantic recognition text data to be confirmed, obtain the confirmation result and send the confirmation result to the video conference server.

[0172] In addition, the device further includes:

[0173] A permission configuration module 809, configured to determine the permission information corresponding to the speech content in response to a triggering operation of a permission configuration component;

[0174] Wherein, a third sending module 806 is further configured to send the permission information corresponding to the speech content and the speech content to the video conference server together.

[0175] It should be noted that the above explanations in the method embodiments for data display in video conferences can also be applied to the device for data display in video conferences in the embodiments of the present application, and will not be elaborated here.

[0176] According to the device for data display in a video conference in an embodiment of the present application, the video conference terminal device obtains the paraphrased speech content of a target participant in the current video conference, sends the paraphrased speech content to the video conference server, so that the video conference server determines target speech data associated with the paraphrased speech content based on historical video conference data, receives the target speech data sent by the video conference server, and displays the target speech data in a target area on the display interface. In this way, the target participant can paraphrase the meeting content based on the target speech data, thereby greatly improving the accuracy and comprehensiveness of the content paraphrased by the user.

[0177] Figure 9 It is a structural block diagram of a video conference system provided by an embodiment of the present application. As Figure 9 shown, the video conference system includes a video conference server 901 and video conference terminal devices 902. Among them, the number of video conference terminal devices 902 can be multiple. The video conference server 901 is configured to execute the method for data display in a video conference applied to the video conference server in the above embodiments, and the video conference terminal devices 902 are configured to execute the method for data display in a video conference applied to the video conference terminal devices in the above embodiments.

[0178] Next, the interaction process between the video conference server 901 and the video conference terminal devices 902 will be introduced:

[0179] (1) The video conference terminal device 902 obtains the paraphrased speech content of a target participant in the current video conference, and sends the paraphrased speech content to the video conference server 901;

[0180] (2) The video conferencing server 901 receives the content of the paraphrased speech of the target participant in the current video conference sent by the video conferencing terminal device 902, determines the target speech data associated with the content of the paraphrased speech based on the stored historical video conference data, and sends the target speech data to the video conferencing terminal device 902;

[0181] (3) The video conferencing terminal device 902 receives the target speech data sent by the video server 901 and displays the target speech data in the target area on the display interface.

[0182] It should be noted that the above explanations in the method embodiments regarding data display in video conferences can also be applied to the video conferencing system in the embodiments of the present application, and will not be elaborated here.

[0183] According to the video conferencing system of the embodiments of the present application, in a video conference, if a target participant needs to paraphrase the conference content to others, the video conferencing server can determine the target speech data associated with the content of the target participant's paraphrased speech based on the stored historical video conference data, and send the target speech data to the target video conferencing terminal device for the target video conferencing terminal device to display the target speech data. In this way, the target participant can paraphrase the conference content based on the target speech data, thereby greatly improving the accuracy and comprehensiveness of the content paraphrased by the user.

[0184] To implement the above embodiments, the present application provides an electronic device.

[0185] Figure 10 The block diagram of the structure of an electronic device provided by the embodiments of the present application. This electronic device can be a server, a computer, etc. As Figure 10 shown, this electronic device includes:

[0186] A memory 1010 and a processor 1020, a bus 1030 connecting different components (including the memory 1010 and the processor 1020), and the memory 1010 stores executable instructions for the processor 1020; wherein, the processor 10100 is configured to execute the instructions to implement the method for data display in the video conference described in the embodiments of the present disclosure.

[0187] The bus 1030 represents one or more of several types of bus structures, including a memory bus or a memory controller, a peripheral bus, a graphics acceleration port, a processor, or a local bus using any bus structure in multiple bus structures. For example, these architectures include, but are not limited to, Industry Standard Architecture (ISA) bus, Micro Channel Architecture (MAC) bus, Enhanced ISA bus, Video Electronics Standards Association (VESA) local bus, and Peripheral Component Interconnect (PCI) bus.

[0188] The electronic device 1000 typically includes a variety of electronically readable media. These media can be any available media accessible to the electronic device 1000, including volatile and non-volatile media, removable and non-removable media. The memory 1010 may also include computer system readable media in the form of volatile memory, such as random access memory (RAM) 1040 and / or cache memory 1050. The electronic device 1000 may further include other removable / non-removable, volatile / non-volatile computer system storage media. By way of example only, the storage system 1060 can be used for reading and writing on non-removable, non-volatile magnetic media ( Figure 10 not shown, commonly referred to as a "hard disk drive"). Although Figure 10 not shown in, a disk drive for reading and writing on a removable non-volatile disk (such as a "floppy disk") and an optical disk drive for reading and writing on a removable non-volatile optical disk (such as a CD-ROM, DVD-ROM or other optical media) can be provided. In these cases, each drive can be connected to the bus 1030 through one or more data media interfaces. The memory 1010 may include at least one program product having a set (such as at least one) of program modules configured to perform the functions of the embodiments of the present disclosure.

[0189] A program / utility 1080 having a set (at least one) of program modules 1070 can be stored, for example, in the memory 1010. Such program modules 1070 include - but are not limited to - an operating system, one or more application programs, other program modules, and program data. Each or some combination of these examples may include the implementation of a network environment. The program modules 1070 generally perform the functions and / or methods in the embodiments described in the present disclosure.

[0190] The electronic device 1000 can also communicate with one or more external devices 1090 (such as a keyboard, a pointing device, a display 1091, etc.), and can also communicate with one or more devices that enable a user to interact with the electronic device 1000, and / or communicate with any device that enables the electronic device 1000 to communicate with one or more other computing devices (such as a network card, a modem, etc.). Such communication can be carried out through an input / output (I / O) interface 1092. Moreover, the electronic device 1000 can also communicate with one or more networks (such as a local area network (LAN), a wide area network (WAN), and / or a public network, such as the Internet) through a network adapter 1093. As shown in the figure, the network adapter 1093 communicates with other modules of the electronic device 1000 through a bus 1030. It should be understood that although not shown in the figure, other hardware and / or software modules can be used in combination with the electronic device 1000, including but not limited to: microcode, device drivers, redundant processing units, external disk drive arrays, RAID systems, tape drives, and data backup storage systems, etc.

[0191] The processor 1020 executes various functional applications and data processing by running programs stored in the memory 1010.

[0192] It should be noted that for the implementation process and technical principle of the electronic device in this embodiment, refer to the foregoing explanation of the method for data display in the video conference of the present disclosure, which will not be elaborated here.

[0193] To implement the above embodiments, the present disclosure also proposes a computer storage medium.

[0194] Wherein, when the instructions in the storage medium are executed by the processor of the server, the server can execute the method for data display in the video conference as described above. Optionally, the computer-readable storage medium can be a ROM, a random access memory (RAM), a CD-ROM, a magnetic tape, a floppy disk, and an optical data storage device, etc.

[0195] In the description of this specification, the description with reference to terms such as "one embodiment", "some embodiments", "example", "specific example", or "some examples" means that the specific features, structures, materials, or characteristics described in connection with the embodiment or example are included in at least one embodiment or example of the present application. In this specification, the schematic representations of the above terms do not necessarily refer to the same embodiment or example. Moreover, the specific features, structures, materials, or characteristics described can be combined in a suitable manner in any one or more embodiments or examples. In addition, without contradiction, those skilled in the art can combine and combine the different embodiments or examples described in this specification and the features of different embodiments or examples.

[0196] In addition, the terms "first" and "second" are used for descriptive purposes only and should not be construed as indicating or implying relative importance or implicitly specifying the quantity of the technical features indicated. Thus, features defined with "first" and "second" may explicitly or implicitly include at least one such feature. In the description of the present application, the meaning of "a plurality" is at least two, such as two, three, etc., unless otherwise specifically defined.

[0197] Any process or method description represented in a flowchart or otherwise described herein may be understood to represent a module, segment, or portion of code including one or more executable instructions for implementing a customized logical function or process, and the scope of the preferred embodiments of the present application includes additional implementations in which functions may be executed in a substantially simultaneous manner or in a reverse order according to the functions involved, rather than in the order shown or discussed, which should be understood by those skilled in the art to which the embodiments of the present application pertain.

[0198] The logic and / or steps represented in a flowchart or otherwise described herein, for example, may be considered as an ordered listing of executable instructions for implementing a logical function and may be specifically implemented in any computer-readable medium for use by or in connection with an instruction execution system, apparatus, or device, such as a computer-based system, a system including a processor, or other systems that can fetch and execute instructions from the instruction execution system, apparatus, or device. As used in this specification, "computer-readable medium" can be any device that can contain, store, communicate, propagate, or transport a program for use by or in connection with an instruction execution system, apparatus, or device. More specific examples (a non-exhaustive list) of the computer-readable medium include the following: an electrical connection portion having one or more wirings (electronic device), a portable computer diskette (magnetic device), a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber device, and a portable compact disc read-only memory (CDROM). Additionally, the computer-readable medium can even be paper or other suitable media on which the program can be printed, as the program can be obtained electronically, for example, by optically scanning the paper or other media, followed by editing, interpretation, or other suitable processing as necessary, and then storing it in a computer memory.

[0199] It should be understood that each part of the present application can be implemented by hardware, software, firmware, or a combination thereof. In the above embodiments, multiple steps or methods can be implemented by software or firmware stored in a memory and executed by a suitable instruction execution system. For example, if implemented by hardware, as in another embodiment, any one of the following techniques known in the art or a combination thereof can be used: discrete logic circuits having logic gate circuits for implementing logic functions on data signals, application-specific integrated circuits having appropriate combinational logic gate circuits, programmable gate arrays (PGAs), field-programmable gate arrays (FPGAs), and the like.

[0200] Those of ordinary skill in the art can understand that all or part of the steps carried by the methods of the above embodiments can be completed by instructing relevant hardware through a program, and the program can be stored in a computer-readable storage medium. When the program is executed, it includes one or a combination of the steps of the method embodiments.

[0201] In addition, in each embodiment of the present application, each functional unit can be integrated in a processing module, can exist physically separately for each unit, or two or more units can be integrated in a module. The above integrated module can be implemented in the form of hardware or in the form of a software functional module. When the integrated module is implemented in the form of a software functional module and sold or used as an independent product, it can also be stored in a computer-readable storage medium.

[0202] Although the embodiments of the present application have been shown and described above, it can be understood that the above embodiments are exemplary and should not be construed as limiting the present application. Those of ordinary skill in the art can make changes, modifications, substitutions, and variations to the above embodiments within the scope of the present application.

Claims

1. A method for displaying data in a video conference, characterized in that: Applicable to video conferencing servers, including: Receiving the target participant's retelling speech content in the current video conference sent by the target video conference terminal device; wherein the target video conference terminal device is the video conference terminal device held by the target participant; the retelling speech content is the target participant's speech content after the retelling operation is triggered; Based on the stored historical video conference data, determining the target speech data associated with the relayed speech content; the historical video conference data includes the speech data of each historical participant in the historical video conference; Sending the target speech data to the target video conferencing terminal device so that the target video conferencing terminal device displays the target speech data; The step of determining target speech data associated with the reported speech content based on the stored historical video conference data includes: Determine the target participants for relaying the speech content; Based on the historical video conference data, determine the target historical video conference attended by the target participant, and the premise for the target participant to relay the content of a certain historical video conference is that the target participant also participated in the historical video conference; Based on the historical video conference data, determining a matching value between the relayed speech content and the speech data of each historical participant in the target historical video conference; According to the matching value, target speech data associated with the reported speech content is determined.

2. The method according to claim 1, characterized in that The step of determining target speech data associated with the reported speech content according to the matching value includes: Determine the authority information in the speech data of each historical participant in the target historical video conference; Determine first target speech data from speech data of each historical participant in the target historical video conference according to the permission information; According to the matching value, target speech data associated with the relayed speech content is determined from the first target speech data.

3. The method according to claim 1, characterized in that The step of determining target speech data associated with the reported speech content according to the matching value includes: The matching value is compared with a preset matching degree threshold, and the speech data of the historical conference participants whose matching value is greater than the matching degree threshold are determined as the target speech data.

4. The method according to claim 1, characterized in that: The method further comprises: Receiving speech content of each participant sent by multiple video conferencing terminal devices corresponding to the current video conference; Performing semantic recognition on the speech content to obtain corresponding semantic recognition text data; Determine the speech content and the semantically recognized text data as speech data corresponding to the conference participant; The speech data of each participant is stored in the historical video conference data.

5. The method according to claim 4, characterized in that The performing semantic recognition on the speech content to obtain corresponding semantic recognition text data includes: Performing semantic recognition on the speech content to obtain corresponding semantic recognition text data to be confirmed; Based on the semantically recognized text data to be confirmed, sending a text data confirmation request to a first video conferencing terminal device held by a corresponding participant; Receive a confirmation result sent by the first video conferencing terminal device, and determine corresponding semantic recognition text data according to the confirmation result.

6. The method according to claim 4, characterized in that The receiving the speech content of each participant sent by multiple video conference terminal devices corresponding to the current video conference includes: Receiving the speech content of each participant and the authority information corresponding to the speech content sent by multiple video conferencing terminal devices corresponding to the current video conference; The step of determining the speech content and the semantically recognized text data as speech data corresponding to the conference participant includes: The speech content, the authority information corresponding to the speech content and the semantic recognition text data are determined as the speech data of the corresponding participant.

7. A method for displaying data in a video conference, characterized in that: Applied to video conferencing terminal equipment, including: Obtaining the retelling speech content of the target participant holding the video conference terminal device in the current video conference; wherein the retelling speech content is the speech content of the target participant after the retelling operation is triggered; The reported speech content is sent to a video conference server; wherein the video conference server is used to determine target speech data associated with the reported speech content based on stored historical video conference data; Receiving target speech data sent by the video server; Displaying the target speech data in a target area of ​​a display interface; The step of determining target speech data associated with the reported speech content based on the stored historical video conference data includes: Determine the target participants corresponding to the relayed speech content; Based on the historical video conference data, determine the target historical video conference attended by the target participant, and the premise for the target participant to relay the content of a certain historical video conference is that the target participant also participated in the historical video conference; Based on the historical video conference data, determining a matching value between the relayed speech content and the speech data of each historical participant in the target historical video conference; According to the matching value, target speech data associated with the reported speech content is determined.

8. The method according to claim 7, characterized in that The step of displaying the target speech data in a target area of ​​the display interface includes: The semantically recognized text data in the target speech data is displayed in the target area.

9. The method according to claim 7, characterized in that: If the number of the target speech data is multiple groups; displaying the target speech data in a target area of ​​the display interface includes: The multiple groups of target speech data are scrolled and displayed in the target area.

10. The method according to claim 7, characterized in that Also includes: In response to a sharing triggering operation of the target speech data, determining conference participants to be shared according to the sharing triggering operation; The target speech data is sent to the video conferencing terminal device held by the participant to be shared.

11. The method according to claim 9, characterized in that The target speech data includes permission information; the sending of the target speech data to the video conference terminal device held by the participant to be shared includes: Determine, according to the permission information, a target sharing participant with sharing permission from among the participants to be shared; The target speech data is sent to the video conferencing terminal device held by the target sharing participant.

12. The method according to claim 7, characterized in that Also includes: Sending the speech content of the target participant in the current video conference to the video conference server; After receiving the speech content, the video conference server performs semantic recognition on the speech content, and stores the obtained semantic recognition text data and the speech content as the speech data of the target participant in the historical video conference data.

13. The method according to claim 12, characterized in that Also includes: Receiving a text data confirmation request sent by the video conferencing server; Determining the semantically recognized text data to be confirmed according to the text data confirmation request, and displaying the semantically recognized text data to be confirmed in a corresponding area of ​​the display interface so that the target conference participant can perform a confirmation operation; In response to the target conference participant's confirmation operation on the semantically recognized text data to be confirmed, a confirmation result is obtained and the confirmation result is sent to the video conference server.

14. The method according to claim 12, characterized in that Also includes: In response to a triggering operation of the permission configuration component, determining permission information corresponding to the speech content; The permission information corresponding to the speech content is sent to the video conference server together with the speech content.

15. A device for displaying data in a video conference, characterized in that: Applicable to video conferencing servers, including: The first receiving module is used to receive the target participant's relayed speech content in the current video conference sent by the target video conference terminal device; wherein the target video conference terminal device is the video conference terminal device held by the target participant; the relayed speech content is the speech content of the target participant after the relay operation is triggered; A first determination module is used to determine target speech data associated with the relayed speech content based on stored historical video conference data; the historical video conference data includes speech data of each historical participant in the historical video conference; A first sending module, configured to send the target speech data to the target video conferencing terminal device, so that the target video conferencing terminal device displays the target speech data; The step of determining target speech data associated with the reported speech content based on the stored historical video conference data includes: Determine the target participants corresponding to the relayed speech content; Based on the historical video conference data, determine the target historical video conference attended by the target participant, and the premise for the target participant to relay the content of a certain historical video conference is that the target participant also participated in the historical video conference; Based on the historical video conference data, determining a matching value between the relayed speech content and the speech data of each historical participant in the target historical video conference; According to the matching value, target speech data associated with the reported speech content is determined.

16. A device for displaying data in a video conference, characterized in that: Applied to video conferencing terminal equipment, including: A first acquisition module is used to acquire the retelling speech content of the target participant holding the video conference terminal device in the current video conference; wherein the retelling speech content is the speech content of the target participant after the retelling operation is triggered; A second sending module is used to send the reported speech content to a video conference server; wherein the video conference server is used to determine target speech data associated with the reported speech content based on the stored historical video conference data; A second receiving module, used for receiving the target speech data sent by the video server; A display module, used to display the target speech data in a target area of ​​a display interface; The step of determining target speech data associated with the reported speech content based on the stored historical video conference data includes: Determine the target participants corresponding to the relayed speech content; Based on the historical video conference data, determine the target historical video conference attended by the target participant, and the premise for the target participant to relay the content of a certain historical video conference is that the target participant also participated in the historical video conference; Based on the historical video conference data, determining a matching value between the relayed speech content and the speech data of each historical participant in the target historical video conference; According to the matching value, target speech data associated with the reported speech content is determined.

17. A video conferencing system, characterized in that: include: A video conferencing server, configured to execute the method according to any one of claims 1 to 6; The video conferencing terminal device is configured to execute the method described in any one of claims 7 to 14.

18. An electronic device, characterized in that: include: processor; a memory for storing executable instructions for the processor; The processor is configured to execute the instructions to implement the method according to any one of claims 1 to 6, and / or to implement the method according to any one of claims 7 to 14.

19. A computer-readable storage medium, characterized in that: When the instructions in the computer-readable storage medium are executed by a processor of an electronic device, the electronic device is enabled to execute the method as described in any one of claims 1 to 6, and / or implement the method as described in any one of claims 7 to 14.

Citation Information

Patent Citations

  • Online conference method and device, electronic equipment and readable storage medium

    CN115623133A