Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

37 results about "Vocal tract" patented technology

The vocal tract is the cavity in human beings and in animals where the sound produced at the sound source (larynx in mammals; syrinx in birds) is filtered. In birds it consists of the trachea, the syrinx, the oral cavity, the upper part of the esophagus, and the beak. In mammals it consists of the laryngeal cavity, the pharynx, the oral cavity, and the nasal cavity.

Chinese learner-oriented tone evaluation and improvement method

PendingCN121415801ASpeech recognitionTime domainVocal tract
The invention relates to the technical field of speech recognition, in particular to a Chinese learner-oriented tone evaluation and improvement method, which comprises the following steps of: extracting a fundamental frequency F0 curve and sound channel parameters of a speech signal, constructing a turbid and clear adaptive fusion model, and fusing to generate fundamental frequency related characteristics and fundamental frequency unrelated characteristics; constructing a tone error corpus, carrying out tone classification labeling, generating a multi-level tone feature set through a turbid and clear adaptive fusion model, aligning a voice signal with a reference voice time domain, and extracting and decoupling a tone shape feature vector and a tone domain feature vector; and based on the decoupled feature vectors, establishing a dual-channel evaluation path, hierarchically calculating the tone type distance and the tone domain distance of the tones, carrying out weighted fusion, generating a final evaluation score, and finally generating feedback information for the Chinese learner. According to the method, a complete method from feature decoupling to two-channel evaluation is constructed, the pronunciation problem of the learner is quantitatively diagnosed, and the pertinence and efficiency of Chinese tone learning are improved.
Owner:GUANGXI UNIV

Semi-closed sound channel training device for high pitch training and use method thereof

The invention aims to provide the semi-closed sound channel training device for high pitch training, which is applied to semi-closed sound channel high pitch range training and can adjust the training airflow resistance, and the use method of the semi-closed sound channel training device. The sealing device comprises a sealing cover connected with the bottle body and a gas inlet pipe penetrating through the sealing cover and used for gas interaction. The air inlet end of the air inlet pipe is connected with a flow guide cover used for being tightly attached to the nose of a user, and an air flow cavity used for guiding air flow is formed in the flow guide cover. The flow guide cover covering the face of a user is adopted, the nasal cavity is communicated with the closed space in the bottle body through the air inlet pipe and the communicating pipe, the user can form airflow resistance by closing the mouth, using the nose to discharge air and produce sound and limiting the air flowing speed through the air pressure adjusting piece, and then the purpose of high-pitch semi-closed sound track training is achieved. The sound channel training device is applied to the technical field of sound channel training devices.
Owner:广州市南沙区唱星文化传媒咨询中心(个体工商户)

Digital human generation system and method based on language model

PendingCN121811908ASpeech analysisTongue tipSpeech sounds
The invention discloses a digital human generation system and method based on a language model, and relates to the technical field of artificial intelligence. A complete digital human pronunciation modeling path from voice signal acquisition, formant frequency extraction, deviation analysis and pronunciation stability evaluation to motion compensation control and three-dimensional animation fusion is realized. Compared with the existing mode of driving the digital human pattern only based on energy envelope or mouth shape classification, the method has the advantages that the multi-frame formant frequency of the user voice is extracted and processed, so that the digital human can generate dynamic lingual surface and tongue tip actions according to the sound channel change in the real pronunciation process of the user; therefore, the physical correspondence and linguistic consistency of the pronunciation actions of the digital human are improved. According to the method, the direct mapping between the pronunciation characteristics and the three-dimensional tongue actions is realized, and the authenticity of the pronunciation animation of the digital human and the guidance in a teaching scene are remarkably improved.
Owner:SUZHOU LVHUA TECH CO LTD

Representation of the speech apparatus in an articulatory feature space

PCT designated stageWO2026124940A1TracheaeSensorsData streamCharacteristic space
Techniques for processing measurement data streams for a person's speech apparatus are disclosed. The measurement data streams can be recorded by different sensor modalities, such as audio recordings or articulatory measurements, for example, which have an observable which quantifies a characteristic property of the vocal tract. The measurement data streams can be processed by means of a machine-learned model in order to obtain a temporal sequence of feature vectors in an articulatory feature space. These specific feature vectors enable a variety of applications, such as articulation training, therapeutic applications in the event of speech restrictions, for instance after a stroke, vocal training or speech synthesis.
Owner:ALTAVO GMBH

Vocal music feature decoupling identification method and system based on mutual information minimization

The invention discloses a nasal sound recognition method based on acoustic feature decoupling, which is characterized in that a double-flow decoupling deep neural network is constructed, a mutual information minimization adversarial training strategy is introduced, and sound source features (vocal cord vibration) and sound channel features (oral cavity and nasal cavity adjustment) are forcibly separated from a single audio signal. On the basis, the high-resolution characteristic of CQT is used for accurately extracting acoustic characteristics related to the nasal sound, physiological offset correction is carried out in combination with personal acoustic fingerprints, and finally accurate quantification and recognition of different nasal sound states (normal, defect and skill) are achieved, so that an objective and visual feedback basis with artistic style adaptability is provided for vocal music teaching.
Owner:HUAZHONG UNIV OF SCI & TECH

A toolkit to aid in learning the sounds of the French language, a process and system for creating these tools, and a learning method implementing this kit.

ActiveFR3154221B1Stammering correctionTeaching apparatusVocal tractPlace of articulation
A kit of tools to aid in learning the sounds of the French language, designed to be applied to the vocal tract of a human subject, each tool comprising a functional part arranged at the end of a handle and shaped to make contact with a given place of articulation from among a plurality of predetermined places of articulation within the vocal tract and to constrain one or more organs of said vocal tract in a determined spatial configuration, each of said tools being associated with a set of consonants. See Figure 3
Owner:BAUMANN ED

Radar marker

PendingUS20250363992A1SurgeryDiagnostic markersVocal tractTesting Methods
Techniques for characterizing a vocal tract by using radar measurements. A radar marker that is arranged in or on the vocal tract is used.
Owner:ALTAVO GMBH

A robust distance estimation method and system based on normal mode-constrained modes in dual-channel audio in Arctic ice regions

This invention provides a tolerance-based distance estimation method and system for dual-channel acoustic imaging in Arctic ice regions based on normal mode-constrained modes. The method includes: using a normal mode model to simulate and predict the frequency domain sound pressure, phase velocity, group velocity, and mode covariance matrix of the experimental sea area in the Arctic ice region; using Fast Fourier Transform (FFT) to transform the time-domain signal received by the vertical array in the ice region into frequency-domain data; obtaining the sound velocity and depth corresponding to the upper boundary, channel axis, and lower boundary of the Beaufort waveguide based on the dual-channel sound velocity profile of the experimental sea area; obtaining the order index of the Beaufort waveguide modes based on the intersection of the phase velocity dispersion curve with the sound velocity at the upper boundary and channel axis of the Beaufort waveguide; performing mode separation using a matched filtering mode filtering method to extract the modes confined in the Beaufort waveguide; and constructing a tolerance-based distance estimation operator for ice and seabed based on the extracted modes in the Beaufort waveguide to achieve tolerance-based estimation of underwater target distances.
Owner:INST OF ACOUSTICS CHINESE ACAD OF SCI

Multi-channel ultrasonic flowmeter measuring method based on pressure correction and application

The invention relates to a multichannel ultrasonic flowmeter measuring method and application based on pressure correction, and the measuring method comprises the steps: uniformly and annularly arranging a plurality of ultrasonic transducers on the same circumferential section, and collecting signal data of a plurality of groups of sound channels; two pressure sensors are arranged at the front position and the rear position of the circumferential section respectively, and the pressure difference and the average pressure are calculated; correcting the instantaneous flow velocity of each group of sound channels based on the average pressure; determining a flow field model according to the pressure difference, and calculating the weighted average flow velocity of the plurality of groups of sound channels according to the instantaneous flow velocity; correcting a cross-sectional area of the circumferential cross-section based on the average pressure; and calculating the medium flow according to the corrected sectional area of the circumferential section and the weighted average flow velocity. The defect that an ultrasonic flowmeter is not high in measurement precision during measurement of a large-pipe-diameter and high-pressure pipeline is overcome, the measurement accuracy can be remarkably improved, the influence of errors of a single sensor or local flow field abnormity on the overall result is reduced, and stability is good.
Owner:GUANGZHOU AOSONG ELECTRONIC CO LTD

An apparatus and method for encoding an audio signal having a plurality of channels

ActiveEP3471091B1Speech analysisVocal tractAudio frequency
An apparatus for encoding an audio signal having a plurality of channels is provided. The apparatus comprises a downmixer (1010) for down mixing the plurality of channels to obtain a downmix signal. Moreover, the apparatus comprises a residual signal calculator (1020) adapted for calculating a residual signal. Furthermore, the apparatus comprises a phase information calculator (1030) adapted for calculating information on a phase difference between the downmix and the residual signal to obtain phase information. Moreover, the apparatus comprises an output generator (1040) for outputting the phase information.
Owner:FRAUNHOFER GESELLSCHAFT ZUR FORDERUNG DER ANGEWANDTEN FORSCHUNG EV

Touch screen voice wake-up system based on voiceprint recognition

PendingCN121983065Awill not leakImprove feature strengthSpeech analysisVocal tractTouchscreen
The invention discloses a touch screen voice wake-up system based on voiceprint recognition, which relates to the technical field of voice data processing and comprises an intelligent control terminal used for controlling data transmission and information interaction among modules. According to the method, various calculations are carried out on the wake-up voice data input by the user in advance, the sound channel shape, the acoustic structure feature and the feature vector independently belonging to the wake-up voice data of the user are determined, after the touch screen collects the surrounding voice data subsequently, feature emphasis is carried out on the first to-be-analyzed voice data so as to improve the feature intensity, and finally, the first to-be-analyzed voice data is analyzed. And verifying the sound channel shape, the acoustic structure feature and the feature vector of the to-be-verified voice through the sound channel shape, the acoustic structure feature and the feature vector of the wake-up voice data, determining whether the first to-be-analyzed voice data is sent by a user recording the wake-up voice data, if so, waking up the touch screen, and if not, not waking up the touch screen. And the privacy of the user is prevented from being leaked.
Owner:HEYUAN LIANTENG WULIAN TECH CO LTD

A voice detection method

The application provides a voice detection method, comprising: obtaining a target voice, preprocessing the target voice, the preprocessing comprising pre-emphasis, framing and windowing; determining a first vocal tract feature, a first sound source wave feature and a plurality of first correlation features of the preprocessed target voice; determining the first principal component feature based on the first vocal tract feature, the first sound source wave feature and the plurality of first correlation features; inputting the first principal component feature into a trained classifier, and outputting a classification result, the classification result being a fake voice or a natural voice. The application utilizes the trace information left by the fake voice at the fundamental frequency, and utilizes the difference between the fake voice and the natural voice in the sound source and the vocal tract feature to realize fake voice detection. The method of principal component analysis is used to screen the sound source and the vocal tract feature respectively, the principal component with high correlation is selected as the feature, the feature dimension and the redundant feature are reduced, and the generalization ability and the efficiency of the model are improved.
Owner:INST OF ACOUSTICS CHINESE ACAD OF SCI +1

Semi-occluded vocal tract exercise device with nonlinear-path and adjustable resistance

An SOVT (semi-occluded vocal tract) exercise device is disclosed, consisting of a compact tubular body with an internal nonlinear airflow path that significantly extends the effective path length of exhaled air. In preferred embodiments a helical partition or insert inside the tube creates a tortuous channel, providing enhanced back-pressure and inertance for vocal training. The device further includes one or more adjustable vent apertures along the tube's wall, which can be selectively opened or closed (for example, by finger, plug, valve, or sleeve) to vary the resistance and / or effective tube length in real time, Certain embodiments incorporate a deformable or one-way valve element at the distal end of the channel that flexes or oscillates under airflow, introducing a pulsating resistance for additional therapeutic benefit. The modular design facilitates easy disassembly for cleaning, interchangeability of inserts or end elements, and portability for voice therapy, respiratory training, or musical applications.
Owner:LUNDQUIST JOSEPH PATRICK

A method for diagnosing early stage motor neuron disease

The application discloses a kind of early motor neuron disease diagnostic methods, belong to medical auxiliary technical field, specifically include: receiving single task and limb rhythm double task under the condition of knocking voice signal sequence and facial infrared thermography video stream data, after time stamp alignment, division analysis section is according to pre-set vocalization unit, extract sound channel micro-disturbance fluctuation sequence and lip movement trajectory curve from each analysis section, carry out instantaneous phase synchronization analysis and obtain coupling stability parameter value, compare the coupling stability parameter value of two kinds of task conditions under the same vocalization grade, calculate stability attenuation ratio value, the attenuation ratio value of different vocalization grade is composed of attenuation feature vector, input pre-constructed attenuation feature discriminant space, output classification mark by pre-division area in discriminant space, the application is thus realized to the analysis and use of sound channel micro-disturbance and lip movement coupling stability attenuation feature under double task load.
Owner:FUJIAN PROVINCIAL HOSPITAL

A speech enhancement method and system based on neural homomorphic synthesis and phase estimation

This invention discloses a speech enhancement method and system based on neural homomorphic synthesis and phase estimation, comprising the following steps: Step S1: Constructing a homomorphic filtering module to receive noisy speech signals, process the signals, and output noisy speech features, wherein the noisy speech features include at least phase information, excitation information, and vocal tract information; Step S2: Constructing an enhancement module to receive the noisy speech features, process the signals, and output enhanced phase information, excitation information, and vocal tract information; Step S3: Constructing a post-processing module to synthesize the enhanced phase information, excitation information, and vocal tract information, and output the enhanced speech signal. This invention implements a neural network homomorphic filter to achieve more accurate separation of excitation and vocal tract; simultaneously, this invention specifically sets up a phase estimation module for phase information, utilizing complex spectral loss and anti-winding loss to enhance phase recovery capability.
Owner:HANGZHOU DIANZI UNIV

Semi-closed sound channel training device

ActiveCN224217169Uprevent outflowBest training resistanceMusicPhysical medicine and rehabilitationVocal tract
The utility model aims to provide the semi-closed sound track training device which is small in size, convenient to carry in daily life and capable of adjusting training resistance. The device comprises an air channel block and a sliding block which is connected with the air channel block in a sliding mode and used for adjusting the air flow, an air cavity allowing the air flow to pass through is formed in the air channel block, one end of the air channel block is provided with a mouth holding part communicated with the air cavity, and the other end of the air channel block is provided with a sliding groove communicated with the air cavity. A sliding groove is formed in the upper surface of the air channel block, a clamping block in sliding fit with the sliding groove is formed on the bottom face of the sliding block, the bottom face of the sliding block is tightly attached to the upper surface of the air channel block, two guide blocks are arranged on the upper surface of the air cavity, and each guide block is in sliding fit with the sliding block. The sound track training device is applied to the technical field of sound track training devices.
Owner:GUANGZHOU BALANCE VOICE TECHNOLOGY CO LTD

Acoustic product sound port protection structure

ActiveCN224459931Ureduce lossesReduces the risk of direct impact to the tarpVocal tractSound wave
This utility model discloses a sound port protection structure for an acoustic product, comprising an acoustic product, a shell, a soft rubber sleeve, and a waterproof cloth. The opening of the shell is equipped with a short vertical baffle, a high vertical baffle, and a horizontal baffle, forming staggered upper and lower sound channels. This utility model has a reasonable structural design. The soft rubber sleeve covers the acoustic product, providing cushioning and shock absorption, thus offering initial protection. Utilizing the properties of the waterproof cloth, it allows sound to pass through while blocking liquids and particles, providing excellent waterproof and dustproof effects. Simultaneously, the structural cooperation of the short, high, and horizontal baffles forms a staggered labyrinthine structure of upper, middle, and lower sound channels. The tortuous and staggered sound channel paths effectively block directly intruding rainwater, washing water jets, and larger dust particles, greatly reducing the risk of high-pressure water directly impacting the waterproof cloth, while sound waves can easily propagate through the tortuous sound channel paths with minimal sound loss, ensuring acoustic performance.
Owner:KINGSTATE ELECTRONICS DONGGUAN CO LTD

Cover body for semi-closed sound channel training

ActiveCN224232261UMusicVocal tractBottle
The utility model aims to provide the cover body for semi-closed sound track training, which is simple in structure, convenient to use in daily life and adjustable in resistance. The bottle comprises a bottle body, a cover body in threaded fit with the bottle body, a first air pipe arranged on the upper surface of the cover body and a second air pipe arranged on one side of the first air pipe, an extension pipe used for extending into the bottle body is formed on the lower portion of the cover body, the extension pipe is communicated with the first air pipe, and the second air pipe is communicated with the first air pipe. The first air pipe and the second air pipe are each provided with a pipe plug used for blocking a pipeline, an open hole is formed in the bottom of the cover body, and the open hole is communicated with the lower end of the second air pipe. The semi-closed sound channel training device is applied to the technical field of semi-closed sound channel training devices.
Owner:GUANGZHOU BALANCE VOICE TECHNOLOGY CO LTD

Ear-worn hearing device with physiological or activity sensor

An ear-worn hearing device is disclosed and includes a body portion with a sound-producing transducer acoustically coupled to a sound passage of a nozzle. a resilient portion protrudes from a side of the body portion and includes a physiological or activity sensor coupled to a flexible portion of a flex harness. The resilient portion at least partially covers a portion of the flex harness without impeding operation of the sensor, wherein the flexible portion and the sensor are flexible toward and away from the body portion upon depression and release of the resilient portion.
Owner:KNOWLES ELECTRONICS LLC

Measuring vocal tract movements of a person by means of radio waves

PCT designated stageWO2026176126A1Ultra wideband antennasVocal tract
The invention relates to a device for measuring vocal tract movements of a person. The device comprises at least one flexible carrier structure (180) which is designed to be adhesively bonded to a skin surface in a person's vocal tract region. Two ultra-wideband antennas (81, 86), which are differently oriented, are also provided.
Owner:ALTAVO GMBH

Device to support verbal communication in tracheotomized, intubated, or laryngectomized patients

ActiveDE102024116440B3Tracheal tubesStammering correctionNasal prongsPhysical medicine and rehabilitation
The invention relates to a device for supporting verbal communication in tracheotomized, intubated, or laryngectomized patients. The device comprises the following elements: Means for generating an airflow; and a nozzle applicator connected via a supply tube to the means for generating the airflow, wherein the nozzle applicator has at least one air outlet nozzle designed such that, when used as intended, the airflow can be introduced through the open mouth into a vocal tract of the patient and the airflow exiting the air outlet nozzle produces white noise, characterized in that the supply tube is a nasal cannula with caudally oriented nasal nozzles and the nozzle applicator is connected to at least one of the nasal nozzles.
Owner:EC EICHHORST CONSULTING GMBH

Eartips with an inset silicone foam section

PendingUS20260255095A1Vocal tractSilicon rubber
A deformable eartip comprising: a monolithic silicone rubber eartip body comprising: an annular inner body defining a sound channel through the deformable eartip, and an annular outer flange integrally formed with and surrounding the annular inner body in a spaced apart relationship with the annular inner body; a closed-cell silicone foam section formed on and completely surrounding a portion of an outer surface of the annular inner body and extending to the outer flange thereby filling in space between the annular inner body and outer flange; a deflection zone formed between the annular outer flange and the inner wall; and an annular rigid frame coupled to the annular inner body and defining a central frame opening formed through the frame that is aligned with the sound channel formed through the annular inner body.
Owner:APPLE INC

DEVICE FOR MEASURING A PERSON'S VOCAL TRACT MOVEMENTS USING RADIO WAVES

UndeterminedDE102025106933A1Ultra wideband antennasVocal tract
A device for measuring vocal tract movements of a person is disclosed. This device comprises at least one flexible support structure (180) designed to be adhered to a skin surface in the vocal tract region of the person. It also includes two ultra-wideband antennas (81, 86) oriented differently.
Owner:ALTAVO GMBH +1

Semi-closed sound track training bottle convenient for adjusting resistance

ActiveCN224207328UEasy to adjust, install and put inSimplify daily disassembly and maintenance processMedical devicesMuscle exercising devicesClassical mechanicsVocal tract
The utility model aims to provide the semi-closed sound track training bottle convenient to adjust the resistance, which is convenient to use in daily life and flexible to adjust the air pressure resistance. The bottle comprises a bottle body, wherein a connecting port is formed in the top of the bottle body; the cover body is in locking fit with the connecting port, the cover body is provided with a first pipeline and a second pipeline, one end of the first pipeline penetrates through the cover body and extends into the bottom of the bottle body, one end of the second pipeline is communicated with the cover body, a reserved cavity is formed in the middle of the second pipeline, and the reserved cavity is communicated with the first pipeline. The reserved cavity is rotationally connected with an adjusting piece used for controlling the air flow. And each plug piece is correspondingly and detachably matched with the extending section of the first pipeline or the extending section of the second pipeline. The semi-closed sound track training bottle is applied to the technical field of semi-closed sound track training bottles.
Owner:GUANGZHOU BALANCE VOICE TECHNOLOGY CO LTD

Ocean sound channel feature recognition method based on improved Faster R-CNN

The invention belongs to the technical field of ocean acoustics, and particularly relates to an ocean sound channel feature recognition method based on improved Faster R-CNN, which comprises the following steps: step 1, sound velocity data preprocessing and imaging; step 2, an improved Faster R-CNN algorithm adopts a Res-Net 101 network to extract image features of the input image, and an image feature map is obtained; step 3, selecting candidate regions through a region suggestion network; step 4, performing detection frame screening by a non-maximum suppression method; the method comprises the following steps: introducing a Soft-NMS strategy into an improved Faster R-CNN algorithm, and screening out a detection frame with relatively high confidence coefficient; 5, after the sizes of the candidate regions are improved to be uniform, the sizes of the candidate regions output by the RPN network are fixed, and the candidate regions are directly input to a full connection layer for classification and regression prediction; and processing the vectors of the full connection layer through a Soft-max classifier, and selecting the type corresponding to the maximum probability as the classification of the target. According to the method, the recognition accuracy and the space-time generalization capability are improved, and the sound channel feature recognition capability in a complex marine environment is enhanced.
Owner:HARBIN ENG UNIV

Voiceprint recognition identity verification method and system

InactiveCN122067525ASpeech analysisThroatSoft tissue deformation
The invention relates to the technical field of voice recognition, and discloses a voiceprint recognition identity verification method and system, and the method comprises the steps: collecting an original sound wave signal, and extracting a multi-dimensional acoustic feature set; constructing an elastic stress correlation model by combining the motion tremor frequency of the laryngeal muscles in the original sound wave signal, and generating a dynamic deformation track of the laryngeal organ; sending a random acoustic interference instruction to the user to trigger non-autonomous stress response of the laryngeal muscle, and synchronously collecting sound wave conduction phase offset; reversely analyzing sound channel soft tissue deformation amplitude according to the sound wave conduction phase offset, matching biomechanical transmission characteristics in the dynamic deformation track of the throat organ, and generating a dynamic physiological consistency coefficient; and when the dynamic physiological consistency coefficient and a preset physiological reference value meet an organ movement cooperative constraint condition, outputting an identity verification passing signal. According to the invention, the problem that voiceprint recognition identity verification cannot effectively defend high-fidelity recording playback can be solved.
Owner:FUJIAN JUNNUO SCI & TECH ACHIEVEMENTS TRANSFORMATION SERVICE CO LTD

Vocalization bottle

ActiveCN310177589SVocal tractMechanical engineering
1. The name of the design product: vocal training bottle. 2. The use of the design product: for semi-closed sound channel training. 3. The design points of the design product: in shape. 4. The picture or photo that best shows the design points: design 1 perspective view. 5. Design 1 is designated as the basic design. 6. Other circumstances that need to be explained: design 2 is the shape after adding a mouthpiece to design 1.
Owner:施东雨