Meeting Role Separation Using Sound Source Angle Recognition
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing role separation systems based on voiceprint identification require offline data accumulation, making real-time separation difficult and impacting user experience.
Innovation Solution
Utilizing sound source angle data to identify and separate roles in real-time through sound source localization and identification methods, including eigenvalue decomposition and voice activity detection.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If voiceprint identification is used for role separation, then identification accuracy is improved, but real-time processing capability deteriorates due to offline data accumulation requirements
Solution Approach 1:
The patent changes the identification parameter from voiceprint (requiring time-domain signal accumulation) to sound source angle (spatial parameter). By using sound source angle data from microphone arrays, the system achieves real-time role separation without requiring offline data accumulation, thus resolving the contradiction between identification accuracy and real-time processing capability
Solution Approach 2:
The patent replaces the traditional voiceprint identification mechanism with sound source localization technology. Instead of analyzing voice characteristics over time, the system uses spatial information from multiple microphones to identify roles in real-time, substituting a different physical measurement approach to achieve both accuracy and real-time performance
2Measurement precision
If offline speech data is used for role separation, then identification accuracy is improved, but system complexity and data processing time increase
Solution Approach 1:
The patent performs preliminary action by pre-configuring the sound source angle calculation framework and microphone array parameters. The system prepares the spatial analysis infrastructure in advance, enabling rapid real-time role separation without time-consuming offline data processing, thus reducing data processing time while maintaining accuracy
3Measurement precision
If voiceprint identification is implemented, then role separation accuracy is improved, but device complexity and computational requirements increase
Solution Approach 1:
The patent extracts only the essential spatial information (sound source angle) from the complex audio signal processing task. By focusing solely on directional information from microphone arrays rather than comprehensive voiceprint analysis, the system achieves role separation accuracy with reduced computational complexity and lower system requirements
Data Source
AI summary
A role separation method, a meeting summary recording method, a role display method and apparatus, an electronic device, and a computer storage medium, relating to the field of speech processing. The role separation method comprises: obtaining sound source angle data corresponding to a speech data frame, acquired by a speech acquisition device, of a role to be separated (S102); on the basis of the sound source angle data, performing identity recognition on the role to be separated to obtain a first identity recognition result of the role to be separated (S104); and separating the role on the basis of the first identity recognition result of the role to be separated (S106). The role is separated in real time, thus making user experience smooth.


