Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

9 results about "Speech identification" patented technology

Speech ID helps protect the authenticity and accuracy of speech by helping the hearing aid to prioritize conversation in difficult listening situations. This feature works to isolate the unique frequencies associated with various letters and words and helps maintain the crispness of conversation. Acuity Directionality.

Voice recognition system and method based on audio-visual dual-mode perception

InactiveCN121583241ASpeech recognitionSpeech identificationSpeech sound
The invention discloses a voice recognition system and method based on audiovisual bimodal perception, and relates to the technical field of voice recognition, and the method comprises the steps: carrying out the feature recursion of an aligned bimodal sequence based on an audiovisual frame package, obtaining an audio high-level representation and a visual high-level representation, and calculating the modal reliability; and fusing the audio high-level representation and the visual high-level representation based on the modal reliability, and evaluating the prefix stability and confidence state through a streaming decoder to obtain a streaming identification information packet. According to the method, the audio high-level representation and the visual high-level representation are fused based on the modal reliability, and the streaming decoder is used for evaluating the prefix stability and the confidence state to obtain the streaming identification information packet, so that the decoding output has the confirmable stable fragment and the credibility representation, and the continuity and the consistency of the speech identification text are improved.
Owner:SHANGHAI JIZHI DIGITAL TECHNOLOGY CO LTD

Method for detecting aircraft air conflict based on semantic parsing of control speech

ActiveUS20250342839A1Speech recognitionAircraft traffic controlEngineeringSpeech identification
A method for detecting aircraft flight conflict based on semantic parsing of control speech is provided. Firstly, the control speech is collected in real time, and the speech is converted into text through the control speech identification system combined with real-time radar data; inputting the identified control text into the control text intention identification model, further analyzing the intention, and judging whether the identified control text intention needs to change the flight state of the aircraft, if not, terminating the detection, and if necessary, extracting control instructions; extracting the key data needed for conflict detection from the identified control text; according to the key data and real-time radar data, performing the conflict detection algorithm to judge whether there is flight conflict; if it exists, a conflict alarm mechanism is triggered.
Owner:CIVIL AVIATION FLIGHT UNIV OF CHINA

A method, apparatus, system and product for synthesized speech identification

This application relates to the field of speech authentication technology, and discloses a method, apparatus, system, and product for identifying synthesized speech. The method includes: constructing a target dataset based on real speech data and synthesized speech data from multiple speakers; wherein each speech data has a corresponding speaker label; constructing an authentication model, including a feature extraction module, a classification module, and a judgment module; training the authentication model using the target dataset; after the model training is completed, processing the target speech data through the authentication model to obtain an authentication result; the authentication result is used to indicate whether the target speech data is synthesized speech. This method can improve the accuracy of synthesized speech authentication and effectively recognize highly realistic synthesized speech from speakers.
Owner:CHINA TELECOM ARTIFICIAL INTELLIGENCE TECHNOLOGY (BEIJING) CO LTD

Synthesized speech identification method, apparatus and system, storage medium, and device

PCT designated stageWO2026123823A1Speech analysisFeature extractionData set
The present application discloses a synthesized speech identification method, apparatus and system, a storage medium, and a device. The method comprises: constructing a target data set on the basis of real speech data and synthesized speech data of multiple speakers, wherein each piece of speech data has a corresponding speaker label; constructing an identification model, wherein the identification model comprises a feature extraction module, a classification module, and a determination module; using the target data set to train the identification model; and after the model training is completed, processing target speech data by means of the identification model to obtain an identification result, wherein the identification result is used for indicating whether the target speech data is a synthesized speech.
Owner:CHINA TELECOM ARTIFICIAL INTELLIGENCE TECHNOLOGY (BEIJING) CO LTD

Method for detecting aircraft air conflict based on semantic parsing of control speech

ActiveUS12537007B2Speech recognitionAircraft traffic controlEngineeringSpeech identification
A method for detecting aircraft flight conflict based on semantic parsing of control speech is provided. Firstly, the control speech is collected in real time, and the speech is converted into text through the control speech identification system combined with real-time radar data; inputting the identified control text into the control text intention identification model, further analyzing the intention, and judging whether the identified control text intention needs to change the flight state of the aircraft, if not, terminating the detection, and if necessary, extracting control instructions; extracting the key data needed for conflict detection from the identified control text; according to the key data and real-time radar data, performing the conflict detection algorithm to judge whether there is flight conflict; if it exists, a conflict alarm mechanism is triggered.
Owner:CIVIL AVIATION FLIGHT UNIV OF CHINA

Large model-based multi-feature voice identification method, apparatus and device, and medium

The invention relates to the technical field of voice semantics, and discloses a multi-feature voice identification method based on a large model, and the method comprises the steps: generating a voice fusion feature according to an initial semantic feature and an initial voice feature; and generating a voice identification result of the initial voice information based on a preset voice identification decision model and the voice fusion feature. Through the above mode, the semantic feature and the voice feature are fused to generate the voice fusion feature, and the semantic information and the acoustic information of the voice are comprehensively considered. The voice identification result generated based on the preset voice identification decision model and the voice fusion feature can adapt to voice data of different languages, different speaking styles and different background environments, and the accuracy of identifying the voice information by the voice identification system is improved in the business fields of financial science and technology, medical treatment, health, pension and the like.
Owner:PING AN TECH (SHENZHEN) CO LTD

A dialect speech recognition method and system based on pre-training and fine-tuning

ActiveCN118298814BSpeech recognitionHigh level techniquesData setSpeech identification
The application discloses a dialect speech recognition method and system based on pre-training and fine-tuning, relates to the technical field of speech recognition, and comprises the following steps: obtaining to-be-recognized speech data, obtaining to-be-recognized speech identification information according to the to-be-recognized speech data, obtaining a dialect speech recognition model according to a speech recognition data set, and obtaining speech recognition information based on the dialect speech recognition model according to the to-be-recognized speech identification information. The application improves the accuracy of speech recognition through the to-be-recognized speech identification information, improves the accuracy and reliability of data through an audio data quality index, provides a data basis for a speech recognition model, trains an existing speech recognition model through data set standard segment information, obtains a dialect speech recognition model, screens possible speech recognition information according to a speech matching index, obtains speech recognition fuzzy correction information, and improves the generalization ability of the model and the accuracy of speech recognition.
Owner:HAINAN UNIV +1

Hatred speech identification method and system

The invention relates to the technical field of natural language processing, and discloses a hatred speech recognition method and system.The system finely adjusts a plurality of models under various prompt configurations, hyper-parameters and input templates, reinforcement learning optimization based on human feedback is executed on the models, and finally reasoning results of the models are fused through a voting mechanism, so that a hatred speech recognition result is obtained. The final prediction output with higher stability and higher robustness is formed, and the adaptive capacity of the system under different input styles and expression modes is remarkably improved. According to the system, aiming at a hatred speech recognition problem in Chinese social media, fine-grained structure extraction, human feedback learning, reward modeling and reinforcement learning are organically combined for the first time, a hatred recognition task is divided into two stages of structure recognition and attitude judgment, and reward function modeling and near-end strategy optimization are introduced, so that the recognition accuracy of the hatred speech is improved. According to the method, the recognition capability of the model on the hidden hatred intention in complex speech can be remarkably improved.
Owner:YUNNAN UNIV

Speech identification and extraction from noise using extended high frequency information

Improved systems and methods are provided herein for extracting target speech from audio signals that can contain masking speech or other unwanted noise content. These systems and methods include detection of target speech in an input signal by detecting elevated frequency content in the signal above a threshold frequency. Portions of the signal determined to contain such elevated high frequency content are then used to generate audio filters to extract target speech from subsequently-obtained audio signals. This can include performing non-negative matrix factorization to determine a set of basis vectors to represent noise content in the spectral domain and then using the set of basis vectors to decompose subsequently-obtained audio signals into noise signals that can then be removed from the audio signals.
Owner:THE BOARD OF TRUSTEES OF THE UNIV OF ILLINOIS