Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

5 results about "Consonant" patented technology

In articulatory phonetics, a consonant is a speech sound that is articulated with complete or partial closure of the vocal tract. Examples are [p], pronounced with the lips; [t], pronounced with the front of the tongue; [k], pronounced with the back of the tongue; [h], pronounced in the throat; [f] and [s], pronounced by forcing air through a narrow channel (fricatives); and [m] and [n], which have air flowing through the nose (nasals). Contrasting with consonants are vowels.

An AI algorithm-based optimized sound recognition acceleration method

PendingCN122116878ASpeech recognitionAlgorithmConsonant
The application discloses an AI algorithm optimization-based sound recognition acceleration method, and particularly relates to the technical field of speech recognition, and comprises instantaneous audio frame composite feature extraction, dynamic inference configuration strategy generation, instantaneous reconstruction inference network construction, accelerated forward inference and decoding lattice generation, and optimal recognition text sequence output. The application dynamically specifies a macroscopic calculation path for a pickup input signal by calculating the instantaneous features of audio frames, and establishes shielding rules according to real-time speech speed. During inference, the network is instantaneously reconstructed according to the hierarchical regulation mechanism, shallow paths that skip the encoder layer are used for simple signals such as silence or vowels, and deep paths are called for complex consonants. The strategy double-binds the calculation resources and real-time features of signals, thereby significantly reducing the average calculation load and delay of a terminal device when processing real speech streams without sacrificing the processing capacity of complex signals.
Owner:GUANGZHOU EDGE COMPUTING TECHNOLOGY CO LTD

Method for early recognition of parkinsonian dysarthria based on voiceprint features

ActiveCN121545557BDysarthriaMedical diagnosis
The application discloses a Parkinson disease dysarthria early identification method based on voiceprint features and belongs to the technical field of medical diagnosis. The original speech data of a to-be-identified object is subjected to speech analysis and speech recognition, and an initial feature set related to voiceprints in a pronunciation deviation and a segment is extracted, target features with significant discriminability for Parkinson disease dysarthria identification are screened out, a dysarthria identification result of the to-be-identified object is obtained through a model, accurate vowel and consonant segment extraction and pronunciation deviation identification are realized, phoneme segment accurate cutting is realized by combining the association of a text sequence and speech data, the quantization determination of the pronunciation deviation is realized, the key dimension of the dysarthria identification is reserved through feature extraction, the early signals of Parkinson disease dysarthria are captured, the high-risk groups are marked, the early warning result is output in combination with the joint risk value calculation, the problem that the prodromal symptoms are not obvious is solved, the review interval is dynamically adjusted, and the whole-process management from identification to monitoring is realized.
Owner:SECOND MEDICAL CENT OF CHINESE PLA GENERAL HOSPITAL

Audio data processing methods, devices, and servers

This specification provides a method, apparatus, and server for processing audio data, applicable to the financial field. Based on this method, after receiving target audio data and obtaining the corresponding first data group through speech recognition, the first data group is first split according to a preset splitting rule to obtain multiple character matrices. Then, based on the pinyin data of the characters, the characters in the character matrices are mapped to corresponding letter character combinations to obtain the corresponding first letter character matrix. According to a preset exchange rule, letter character combinations in the first letter character matrix that satisfy the exchange conditions are determined. The initial consonants and / or final vowels of the letter character combinations that satisfy the exchange conditions are then exchanged to obtain the corresponding second letter character matrix. Based on the second letter character matrix, the corresponding second data group is obtained. This method can efficiently hide relevant information carried in the audio data, protecting the data security of that information.
Owner:INDUSTRIAL AND COMMERCIAL BANK OF CHINA

A method and system for generating a multi-dimensional speech training program

ActiveCN121687374BAchieve precise quantitative analysisImprove targetingSpeech trainingSpeech comprehension
The application discloses a kind of multi-dimension speech training plan generation method and system.The method first collects patient age, gender and speech sample, obtains seven-dimensional evaluation parameters including sound pressure, amplitude perturbation, maximum vocalization duration, fundamental frequency perturbation, vowel space area, tone impairment and speech comprehension score.Subsequently, according to the preset logic, these parameters are sequentially determined based on the parameters, and dynamically combine different training modules such as loudness, breath, pitch, vowel, glide, tone and consonant into personalized speech training plan.The application solves the problem of existing technology training scheme solidification through multi-dimensional evaluation and dynamic module matching, significantly improves the individualization degree and rehabilitation effect of speech training.
Owner:BEIJING REHABILITATION HOSPITAL CAPITAL MEDICAL UNIVERSITY(BEIJING WORKERS SANATORIUM)