Large-screen video conference interface control method based on time-space synchronization

By granting tiered speaking permissions to video conference speakers and implementing time and space synchronization, the problem of decreased meeting quality caused by multiple speakers speaking simultaneously has been solved, and efficient multi-screen updates and synchronization have been achieved.

CN121664780APending Publication Date: 2026-03-13北京优易租网络技术有限公司
View PDF 0 Cites 0 Cited by

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
Filing Date
2025-10-20
Publication Date
2026-03-13

AI Technical Summary

Technical Problem

In existing video conferencing, when there are many participants, multiple people speaking at the same time can easily lead to a decline in meeting quality and efficiency, and the meeting can easily fall into chaos when there are interfering factors.

Method used

By acquiring information about meeting participants and the current meeting content, the system identifies speakers and assigns them hierarchical speaking permissions. Microphone access is controlled according to the speaking order and permissions. Interface control commands are generated and synchronized in time and space to achieve multi-screen updates.

Benefits of technology

It effectively avoids disputes over meeting content caused by multiple people speaking at the same time, ensures meeting efficiency, and enables simultaneous updates of multiple screens.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN121664780A_ABST
    Figure CN121664780A_ABST
Patent Text Reader

Abstract

The invention discloses a large-screen video conference interface control method based on time-space synchronization, and relates to the technical field of conference interface control, and the method comprises the steps: obtaining the information of conference participants and the current conference content to determine conference spokesmen, endowing the conference spokesmen with a hierarchical speaking authority, and determining the speaking sequence of the conference spokesmen. And opening microphone permissions of the conference spokesmen according to the graded speaking permissions, respectively obtaining conventional speaking requests and emergency speaking requests of the conference spokesmen and speaking states under the current conference process, endowing the microphone with the open permissions, switching the conference spokesmen, obtaining speaking contents of the conference spokesmen, and sending the speaking contents to the conference spokesmen. And performing content identification to obtain a speaking end feature language, generating an interface control instruction according to the speaking end feature language, recording an operation timestamp and an operation coordinate as time synchronization parameters, executing an interface operation after coordinate conversion in a preset time window of the target control terminal, and realizing synchronous updating of multi-screen pictures.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] This invention relates to the field of conference interface control technology, and in particular to a method for controlling a large-screen video conferencing interface based on spatiotemporal synchronization. Background Technology

[0002] With the continuous development of internet technology, people have increasingly higher demands for timeliness and urgency in problem-solving, leading to the widespread application of video conferencing in daily life. As an important means of remote communication, video conferencing technology has undergone years of development and evolution, from simple audio calls to high-definition video, screen sharing, and multi-party interactive functions. Video conferencing technology continues to break through and innovate, and online large-screen video conferencing, as an important form of video conferencing technology, has seen its interface sharing and control methods continuously optimized and improved with technological advancements.

[0003] With the continuous development of online large-screen video conferencing technology, meeting interface sharing has become one of the important functions of meetings. However, in practical applications, meeting interface sharing faces many challenges. Currently, video conferencing only has two modes: speaking with the microphone on and speaking with the microphone off. When there are many participants, it is easy for multiple people to speak with their microphones on at the same time, affecting the quality of the meeting. Furthermore, when there are other interfering factors, the meeting will fall into chaos, affecting the normal progress of the meeting and thus affecting the efficiency of the meeting. To address the aforementioned technical deficiencies, a solution is proposed. Summary of the Invention

[0004] The purpose of this invention is to: acquire the speech content of the speakers at the meeting, perform content recognition to obtain the speech end feature language, generate interface control instructions based on the speech end feature language, and record the operation timestamp and operation coordinates as time synchronization parameters. The interface operation after coordinate transformation is executed within the preset time window of the target control terminal to realize the synchronous update of multi-screen screens, which not only ensures the meeting efficiency but also avoids disputes caused by multiple people speaking at the same time.

[0005] To achieve the above objectives, the present invention adopts the following technical solution: a large-screen video conferencing interface control method based on spatiotemporal synchronization, comprising the following steps: Step 1: Obtain the personnel information of the meeting participants and the current meeting content; determine the meeting speakers based on the personnel information and the current meeting content; and assign hierarchical speaking permissions to the meeting speakers. Step 2: Determine the speaking order of the speakers based on the current meeting content, and grant microphone permissions to the speakers according to their speaking privileges. On the shared interface of the large-screen video conference, adjust the speaker's screen to the maximum size. Step 3: Obtain the regular speaking requests and speaking status of the conference speakers in the current meeting process, determine their corresponding microphone permissions based on their corresponding hierarchical speaking permissions, grant microphone access permissions, and switch conference speakers; Step 4: Obtain the emergency speaking request from the conference speaker, extract content features based on the emergency speaking request, compare the content features with the current conference content, determine the necessity of the emergency speaking request, and at the same time determine the corresponding microphone permissions based on the corresponding hierarchical speaking permissions, generate microphone access permissions, and switch the conference speaker. Step 5: Obtain the speech content of the conference speaker, perform content recognition to obtain the speech end feature language, and generate interface control instructions based on the speech end feature language. At the same time, record the operation timestamp and operation coordinates as time synchronization parameters. Step Six: Based on the spatiotemporal synchronization parameters, convert the interface control commands into the local coordinates and execution time of the target control terminal of the next conference speaker. Then, send the timestamped interface control commands to the target control terminal synchronously through an encrypted communication channel. Execute the coordinate-converted interface operations within the preset time window of the target control terminal to achieve synchronous updates of multiple screens.

[0006] Furthermore, the specific process for identifying conference speakers and assigning them tiered speaking privileges is as follows: S101. Obtain the personnel information of the meeting participants, including their names, job levels, and functional attributes. At the same time, obtain the current meeting content and determine the main body of the meeting and the speaking process based on the current meeting content. S102. The meeting process includes the meeting opening, discussion of agenda items, summary and confirmation, and meeting end. The meeting roles in the meeting process are determined according to personnel information, and the meeting roles include meeting host, meeting discussant, and meeting observer. S103. Based on the speaking frequency during the meeting process, the meeting host, the meeting participants and the meeting observers are given hierarchical speaking permissions, wherein the meeting host has the first-level speaking permission, and the first-level permission means that the host can speak throughout the entire meeting process. The speaking privileges of the participants in the meeting are at the second level. The second level of privileges is a semi-process speaking state during the meeting process, specifically having the right to speak during the discussion of agenda items and the process of summarizing and confirming. The speaking privileges of meeting observers are at level three, with level one privileges meaning that they cannot speak at any point during the meeting.

[0007] Furthermore, the specific process for determining the speaking order of the speakers based on the current meeting content is as follows: S201. Obtain the current meeting content, determine the meeting progress based on the current meeting content, and select the meeting host and participants as target speakers based on the specific stage of the meeting process. S202. In the meeting opening state, grant the meeting host microphone access, obtain the meeting host's voice content, perform feature recognition based on the voice content, obtain the personal name keywords in the language content, and use them as candidate speakers. S203. Compare and screen the candidate speakers with the target speakers one by one. If there is a candidate speaker among the target speakers, mark him / her as the prepared speaker. After the meeting host finishes speaking, turn off his / her microphone privileges and then turn on the prepared speaker's microphone privileges. S204. During the discussion of agenda items and the summary and confirmation state, the language content of the meeting participants is obtained. Similarly, feature recognition is performed based on the voice content to obtain the keywords of the names in the language content. After verification, the speaking order of the meeting speakers is determined one by one.

[0008] Furthermore, the specific process for obtaining a regular speaking request and switching the conference speaker is as follows: S301. Obtain the regular speaking requests of the conference speakers and the current conference progress. If the current conference progress is the start and end of the conference, obtain the speaking permissions corresponding to the conference speakers. If the conference speakers have level 1 permissions, obtain the speaking status under the current conference progress. If there are no conference speakers speaking, grant them microphone access permissions. If the speaker at the meeting has level 2 or level 3 access, then microphone access will not be granted. S302. If the current meeting process involves discussing agenda items and summarizing and confirming, obtain the speaking permissions corresponding to the meeting speaker. If the meeting speaker has first-level or second-level permissions, obtain the speaking status under the current meeting process. If there is no meeting speaker currently speaking, grant microphone access permission. If there is a speaker currently speaking, obtain the speaker's meeting role. If the meeting role is the meeting host, grant the speaker microphone access permission. If the meeting role is that of a participant in the meeting discussion, microphone access will not be granted. If the speaker's speaking privileges are at level three, then microphone access will not be granted.

[0009] Furthermore, the specific process for obtaining an emergency speaking request and switching the conference speaker is as follows: S401. Obtain the urgent speaking request of the conference speaker, extract features based on the urgent speaking request, obtain feature tone words, judge the tone based on the preset tone word lexicon, and obtain the urgency of the request. S402. Obtain the preset urgency judgment threshold. If the urgency of the request is greater than or equal to the urgency judgment threshold, then the urgent speaking request is a necessary speaking request. S403. The speaking status in the current meeting process. If there is no speaker currently speaking, grant microphone access. If there is a speaker currently speaking, obtain their corresponding speaking permission, i.e. speaking permission in the active state; obtain the speaking permission of the speaker who made the necessary speaking request, i.e. speaking permission in the request state; if the speaking permission in the request state is higher than the speaking permission in the active state, grant microphone access permission. If the request for status speaking permission is lower than the permission to perform status speaking, then microphone access will not be granted.

[0010] Furthermore, the specific process for achieving synchronized updates across multiple screens: S501. Obtain the speaking permissions of the conference speakers, select the conference speaker with the highest speaking permissions as the main speaker, elect the terminal corresponding to the main speaker as the main clock terminal, and use the terminals corresponding to other conference speakers as slave clocks for time synchronization. S502. A two-way time transfer mechanism is used to calculate network transmission delay. The synchronization period is dynamically adjusted according to the network transmission delay calculation result. When network jitter is detected to exceed the threshold, the synchronization interval is automatically shortened to establish a multi-terminal spatiotemporal synchronization framework. S503. By unifying the time base of each conference speaker's terminal via the NTP protocol, the display resolution parameters and spatial coordinate mapping relationships of each terminal are obtained, and a global spatial coordinate system is constructed. Specifically: S5031. Collect the coordinates of the four corner points of the display area of ​​each terminal; S5032. Establish the mapping relationship between the terminal's local coordinates and global coordinates through an affine transformation algorithm; S5033, Store the coordinate transformation matrix of each terminal for real-time calculation; S504. When an interface control command is detected to be generated, record the time synchronization parameters, namely the operation timestamp and operation coordinates.

[0011] In summary, due to the adoption of the above technical solution, the beneficial effects of the present invention are: This spatiotemporal synchronization-based large-screen video conferencing interface control method determines the speakers by acquiring the personnel information of the meeting participants and the current meeting content, and assigns tiered speaking permissions to the speakers to determine the speaking order. Microphone permissions are granted to speakers according to their tiered speaking permissions. The method acquires the speakers' regular speaking requests, emergency speaking requests, and speaking status under the current meeting progress, grants microphone access, switches speakers, acquires their speaking content, performs content recognition to obtain the speech end characteristic language, and generates interface control instructions based on the speech end characteristic language. Simultaneously, it records the operation timestamp and operation coordinates as time synchronization parameters, and executes the coordinate-transformed interface operation within a preset time window on the target control terminal. This achieves synchronous updates of multiple screens, ensuring meeting efficiency while avoiding disputes caused by multiple simultaneous speakers. Attached Figure Description

[0012] Figure 1 A schematic diagram of the overall method flow of the present invention is shown. Detailed Implementation

[0013] The technical solutions of the embodiments of the present invention will be clearly and completely described below with reference to the accompanying drawings. Obviously, the described embodiments are only some embodiments of the present invention, and not all embodiments. Based on the embodiments of the present invention, all other embodiments obtained by those skilled in the art without creative effort are within the scope of protection of the present invention.

[0014] Example: like Figure 1 As shown, the control method for a large-screen video conferencing interface based on spatiotemporal synchronization includes the following steps: Step 1: Obtain the personnel information of the meeting participants and the current meeting content; determine the meeting speakers based on the personnel information and the current meeting content; and assign hierarchical speaking permissions to the meeting speakers. The specific process for determining the speakers at a meeting and assigning them tiered speaking permissions is as follows: S101. Obtain the personnel information of the meeting participants, including their names, job levels, and functional attributes. At the same time, obtain the current meeting content and determine the main body of the meeting and the speaking process based on the current meeting content. S102. The meeting process includes the meeting opening, discussion of agenda items, summary and confirmation, and meeting end. The meeting roles in the meeting process are determined according to personnel information, and the meeting roles include meeting host, meeting discussant, and meeting observer. S103. Based on the speaking frequency during the meeting process, the meeting host, the meeting participants and the meeting observers are given hierarchical speaking permissions, wherein the meeting host has the first-level speaking permission, and the first-level permission means that the host can speak throughout the entire meeting process. The speaking privileges of the participants in the meeting are at the second level. The second level of privileges is a semi-process speaking state during the meeting process, specifically having the right to speak during the discussion of agenda items and the process of summarizing and confirming. The speaking privileges of meeting observers are at level three, with level one privileges meaning that they cannot speak at any point during the meeting.

[0015] Step 2: Determine the speaking order of the speakers based on the current meeting content, and grant microphone permissions to the speakers according to their speaking privileges. On the shared interface of the large-screen video conference, adjust the speaker's screen to the maximum size. The specific process for determining the speaking order of the speakers based on the current meeting content is as follows: S201. Obtain the current meeting content, determine the meeting progress based on the current meeting content, and select the meeting host and participants as target speakers based on the specific stage of the meeting process. S202. In the meeting opening state, grant the meeting host microphone access, obtain the meeting host's voice content, perform feature recognition based on the voice content, obtain the personal name keywords in the language content, and use them as candidate speakers. S203. Compare and screen the candidate speakers with the target speakers one by one. If there is a candidate speaker among the target speakers, mark him / her as the prepared speaker. After the meeting host finishes speaking, turn off his / her microphone privileges and then turn on the prepared speaker's microphone privileges. S204. During the discussion of agenda items and the summary and confirmation state, the language content of the meeting participants is obtained. Similarly, feature recognition is performed based on the voice content to obtain the keywords of the names in the language content. After verification, the speaking order of the meeting speakers is determined one by one.

[0016] Step 3: Obtain the regular speaking requests and speaking status of the conference speakers in the current meeting process, determine their corresponding microphone permissions based on their corresponding hierarchical speaking permissions, grant microphone access permissions, and switch conference speakers; The specific process for obtaining a regular speaking request and switching the conference speaker is as follows: S301. Obtain the regular speaking requests of the conference speakers and the current conference progress. If the current conference progress is the start and end of the conference, obtain the speaking permissions corresponding to the conference speakers. If the conference speakers have level 1 permissions, obtain the speaking status under the current conference progress. If there are no conference speakers speaking, grant them microphone access permissions. If the speaker at the meeting has level 2 or level 3 access, then microphone access will not be granted. S302. If the current meeting process involves discussing agenda items and summarizing and confirming, obtain the speaking permissions corresponding to the meeting speaker. If the meeting speaker has first-level or second-level permissions, obtain the speaking status under the current meeting process. If there is no meeting speaker currently speaking, grant microphone access permission. If there is a speaker currently speaking, obtain the speaker's meeting role. If the meeting role is the meeting host, grant the speaker microphone access permission. If the meeting role is that of a participant in the meeting discussion, microphone access will not be granted. If the speaker's speaking privileges are at level three, then microphone access will not be granted.

[0017] Step 4: Obtain the emergency speaking request from the conference speaker, extract content features based on the emergency speaking request, compare the content features with the current conference content, determine the necessity of the emergency speaking request, and at the same time determine the corresponding microphone permissions based on the corresponding hierarchical speaking permissions, generate microphone access permissions, and switch the conference speaker. The specific process for receiving an emergency speaking request and switching the conference speaker is as follows: S401. Obtain the urgent speaking request of the conference speaker, extract features based on the urgent speaking request, obtain feature tone words, judge the tone based on the preset tone word lexicon, and obtain the urgency of the request. S402. Obtain the preset urgency judgment threshold. If the urgency of the request is greater than or equal to the urgency judgment threshold, then the urgent speaking request is a necessary speaking request. S403. The speaking status in the current meeting process. If there is no speaker currently speaking, grant microphone access. If there is a speaker currently speaking, obtain their corresponding speaking permission, i.e. speaking permission in the active state; obtain the speaking permission of the speaker who made the necessary speaking request, i.e. speaking permission in the request state; if the speaking permission in the request state is higher than the speaking permission in the active state, grant microphone access permission. If the request for status speaking permission is lower than the permission to perform status speaking, then microphone access will not be granted.

[0018] Step 5: Obtain the speech content of the conference speaker, perform content recognition to obtain the speech end feature language, and generate interface control instructions based on the speech end feature language. At the same time, record the operation timestamp and operation coordinates as time synchronization parameters. Step Six: Based on the spatiotemporal synchronization parameters, convert the interface control commands into the local coordinates and execution time of the target control terminal of the next conference speaker. Then, send the timestamped interface control commands to the target control terminal synchronously through an encrypted communication channel. Execute the coordinate-converted interface operations within the preset time window of the target control terminal to achieve synchronous updates of multiple screens.

[0019] The specific process of achieving synchronized updates across multiple screens: S501. Obtain the speaking permissions of the conference speakers, select the conference speaker with the highest speaking permissions as the main speaker, elect the terminal corresponding to the main speaker as the main clock terminal, and use the terminals corresponding to other conference speakers as slave clocks for time synchronization. S502. A two-way time transfer mechanism is used to calculate network transmission delay. The synchronization period is dynamically adjusted according to the network transmission delay calculation result. When network jitter is detected to exceed the threshold, the synchronization interval is automatically shortened to establish a multi-terminal spatiotemporal synchronization framework. S503. By unifying the time base of each conference speaker's terminal via the NTP protocol, the display resolution parameters and spatial coordinate mapping relationships of each terminal are obtained, and a global spatial coordinate system is constructed. Specifically: S5031. Collect the coordinates of the four corner points of the display area of ​​each terminal; S5032. Establish the mapping relationship between the terminal's local coordinates and global coordinates through an affine transformation algorithm; S5033, Store the coordinate transformation matrix of each terminal for real-time calculation; S504. When an interface control command is detected to be generated, record the time synchronization parameters, namely the operation timestamp and operation coordinates.

[0020] This invention identifies speakers by acquiring information about meeting participants and the current meeting content, assigns tiered speaking permissions to speakers, determines the speaking order, and grants microphone access to speakers based on their tiered speaking permissions. It also acquires speakers' regular speaking requests, urgent speaking requests, and speaking status within the current meeting, grants microphone access, switches speakers, acquires their speaking content, performs content recognition to obtain end-of-speech characteristic language, and generates interface control instructions based on this characteristic language. Simultaneously, it records operation timestamps and coordinates as time synchronization parameters, executes the coordinate-transformed interface operations within a preset time window on the target control terminal, and achieves synchronized updates across multiple screens. This ensures meeting efficiency while preventing disputes arising from multiple simultaneous speakers.

[0021] The size of the interval and threshold is set to facilitate comparison. The size of the threshold depends on the amount of sample data and the number of bases set by those skilled in the art for each set of sample data; as long as it does not affect the ratio between the parameter and the quantized value.

[0022] The above formulas are all dimensionless calculations. The formulas are derived from software simulations based on a large amount of collected data to obtain the most recent real-world results. The preset parameters in the formulas are set by those skilled in the art according to the actual situation. The above description is only a preferred embodiment of the present invention, but the scope of protection of the present invention is not limited thereto. Any equivalent substitutions or modifications made by those skilled in the art within the scope of the technology disclosed in the present invention, based on the technical solution and inventive concept of the present invention, should be covered within the scope of protection of the present invention.

Claims

1. A large-screen video conferencing interface control method based on spatiotemporal synchronization, characterized in that, Includes the following steps: Step 1: Obtain the personnel information of the meeting participants and the current meeting content; determine the meeting speakers based on the personnel information and the current meeting content; and assign hierarchical speaking permissions to the meeting speakers. Step 2: Determine the speaking order of the speakers based on the current meeting content, and grant microphone permissions to the speakers according to their speaking privileges. On the shared interface of the large-screen video conference, adjust the speaker's screen to the maximum size. Step 3: Obtain the regular speaking requests and speaking status of the conference speakers in the current meeting process, determine their corresponding microphone permissions based on their corresponding hierarchical speaking permissions, grant microphone access permissions, and switch conference speakers; Step 4: Obtain the emergency speaking request from the conference speaker, extract content features based on the emergency speaking request, compare the content features with the current conference content, determine the necessity of the emergency speaking request, and at the same time determine the corresponding microphone permissions based on the corresponding hierarchical speaking permissions, generate microphone access permissions, and switch the conference speaker. Step 5: Obtain the speech content of the conference speaker, perform content recognition to obtain the speech end feature language, and generate interface control instructions based on the speech end feature language. At the same time, record the operation timestamp and operation coordinates as time synchronization parameters. Step Six: Based on the spatiotemporal synchronization parameters, convert the interface control commands into the local coordinates and execution time of the target control terminal of the next conference speaker. Then, send the timestamped interface control commands to the target control terminal synchronously through an encrypted communication channel. Execute the coordinate-converted interface operations within the preset time window of the target control terminal to achieve synchronous updates of multiple screens.

2. The large-screen video conferencing interface control method based on spatiotemporal synchronization according to claim 1, characterized in that, The specific process for determining the speakers at a meeting and assigning them tiered speaking permissions is as follows: S101. Obtain the personnel information of the meeting participants, including their names, job levels, and functional attributes. At the same time, obtain the current meeting content and determine the main body of the meeting and the speaking process based on the current meeting content. S102. The meeting process includes the meeting opening, discussion of agenda items, summary and confirmation, and meeting end. The meeting roles in the meeting process are determined according to personnel information, and the meeting roles include meeting host, meeting discussant, and meeting observer. S103. Based on the speaking frequency during the meeting process, the meeting host, the meeting participants and the meeting observers are given hierarchical speaking permissions, wherein the meeting host has the first-level speaking permission, and the first-level permission means that the host can speak throughout the entire meeting process. The speaking privileges of the participants in the meeting are at the second level. The second level of privileges is a semi-process speaking state during the meeting process, specifically having the right to speak during the discussion of agenda items and the process of summarizing and confirming. The speaking privileges of meeting observers are at level three, with level one privileges meaning that they cannot speak at any point during the meeting.

3. The large-screen video conferencing interface control method based on spatiotemporal synchronization according to claim 1, characterized in that, The specific process for determining the speaking order of the speakers based on the current meeting content is as follows: S201. Obtain the current meeting content, determine the meeting progress based on the current meeting content, and select the meeting host and participants as target speakers based on the specific stage of the meeting process. S202. In the meeting opening state, grant the meeting host microphone access, obtain the meeting host's voice content, perform feature recognition based on the voice content, obtain the personal name keywords in the language content, and use them as candidate speakers. S203. Compare and screen the candidate speakers with the target speakers one by one. If there is a candidate speaker among the target speakers, mark him / her as the prepared speaker. After the meeting host finishes speaking, turn off his / her microphone privileges and then turn on the prepared speaker's microphone privileges. S204. During the discussion of agenda items and the summary and confirmation state, the language content of the meeting participants is obtained. Similarly, feature recognition is performed based on the voice content to obtain the keywords of the names in the language content. After verification, the speaking order of the meeting speakers is determined one by one.

4. The large-screen video conferencing interface control method based on spatiotemporal synchronization according to claim 1, characterized in that, The specific process for obtaining a regular speaking request and switching the conference speaker is as follows: S301. Obtain the regular speaking requests of the conference speakers and the current conference progress. If the current conference progress is the start and end of the conference, obtain the speaking permissions corresponding to the conference speakers. If the conference speakers have level 1 permissions, obtain the speaking status under the current conference progress. If there are no conference speakers speaking, grant them microphone access permissions. If the speaker at the meeting has level 2 or level 3 access, then microphone access will not be granted. S302. If the current meeting process involves discussing agenda items and summarizing and confirming, obtain the speaking permissions corresponding to the meeting speaker. If the meeting speaker has first-level or second-level permissions, obtain the speaking status under the current meeting process. If there is no meeting speaker currently speaking, grant microphone access permission. If there is a speaker currently speaking, obtain the speaker's meeting role. If the meeting role is the meeting host, grant the speaker microphone access permission. If the meeting role is that of a participant in the meeting discussion, microphone access will not be granted. If the speaker's speaking privileges are at level three, then microphone access will not be granted.

5. The large-screen video conferencing interface control method based on spatiotemporal synchronization according to claim 1, characterized in that, The specific process for receiving an emergency speaking request and switching the conference speaker is as follows: S401. Obtain the urgent speaking request of the conference speaker, extract features based on the urgent speaking request, obtain feature tone words, judge the tone based on the preset tone word lexicon, and obtain the urgency of the request. S402. Obtain the preset urgency judgment threshold. If the urgency of the request is greater than or equal to the urgency judgment threshold, then the urgent speaking request is a necessary speaking request. S403. The speaking status in the current meeting process. If there is no speaker currently speaking, grant microphone access. If there is a speaker currently speaking, obtain their corresponding speaking permission, i.e. speaking permission in the active state; obtain the speaking permission of the speaker who made the necessary speaking request, i.e. speaking permission in the request state; if the speaking permission in the request state is higher than the speaking permission in the active state, grant microphone access permission. If the request for status speaking permission is lower than the permission to perform status speaking, then microphone access will not be granted.

6. The large-screen video conferencing interface control method based on spatiotemporal synchronization according to claim 1, characterized in that, The specific process of achieving synchronized updates across multiple screens: S501. Obtain the speaking permissions of the conference speakers, select the conference speaker with the highest speaking permissions as the main speaker, elect the terminal corresponding to the main speaker as the main clock terminal, and use the terminals corresponding to other conference speakers as slave clocks for time synchronization. S502. A two-way time transfer mechanism is used to calculate network transmission delay. The synchronization period is dynamically adjusted according to the network transmission delay calculation result. When network jitter is detected to exceed the threshold, the synchronization interval is automatically shortened to establish a multi-terminal spatiotemporal synchronization framework. S503. By unifying the time base of each conference speaker's terminal via the NTP protocol, the display resolution parameters and spatial coordinate mapping relationships of each terminal are obtained, and a global spatial coordinate system is constructed. Specifically: S5031. Collect the coordinates of the four corner points of the display area of ​​each terminal; S5032. Establish the mapping relationship between the terminal's local coordinates and global coordinates through an affine transformation algorithm; S5033, Store the coordinate transformation matrix of each terminal for real-time calculation; S504. When an interface control command is detected to be generated, record the time synchronization parameters, namely the operation timestamp and operation coordinates.