Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

8 results about "Vocal quality" patented technology

Factors that determine quality. Production of original tone by the vocal folds (phonation); process of selection, reinforcement and damping of this tone by the resonators (throat, mouth, nose) Types of vocal quality. Breathy, strident, harsh, hypernasality, glottal fry, throaty, glottal attack, hoarse, hyponasality.

Method, device and equipment for voiceprint recognition system to resist attack and medium

PendingCN121963772AImprove migration abilityImprove stabilitySpeech analysisAttacker modelAlgorithm
The invention discloses a voiceprint recognition system attack resisting method and device, equipment and a medium, by constructing a shadow voice sample set and a substitution model, dependence on target speaker data and query capability is reduced, and privacy risk and query overhead are reduced; meanwhile, two types of substitution models are trained based on an information theory method, so that the mobility and stability of the models are improved, and the generalization ability is enhanced; besides, the attacker model adopts a specific architecture and combines an alternate training strategy, so that attack loss and prediction probability difference are optimized, imperceptibility and attack effectiveness are balanced, and tone quality reduction or attack failure is avoided; and finally, through systematic training and optimization, a more uniform evaluation caliber is expected to be provided, and more reliable guidance is provided for engineering landing and risk evaluation.
Owner:GUIZHOU UNIV

An adjustable frequency response microphone with integrated XLR and USB outputs

This invention discloses an adjustable frequency response microphone integrating XLR and USB outputs, comprising a microphone body, a dynamic microphone driver, a passive filter, a pass-through / preamp circuit, an XLR XLR interface, a USB module, and a function control circuit. The passive filter employs an LC circuit structure, supporting three frequency modes: low-frequency cutoff, mid-frequency boost, and flat response, effectively filtering out ambient noise and enhancing vocal quality. The pass-through / preamp circuit offers both pass-through and preamp modes, with adjustable preamp gain to accommodate dynamic microphone drivers of varying sensitivities, and can be externally or internally powered by phantom power. The USB module integrates a Type-C interface and an audio chip, supporting analog-to-digital / digital-to-analog conversion, real-time headphone monitoring, and external power supply. The function control circuit integrates an encoder and a touch panel, enabling adjustments such as volume, mixing, and mute, and enhances the interactive experience with an RGB lighting module. This invention combines the advantages of professional XLR output with convenient USB output, making it suitable for professional recording, live streaming, gaming, and other scenarios.
Owner:ZHAOQING HEJIA ELECTRONICS CO LTD

Innovative karaoke sound system with combined neural network feedback suppression and coded vocal restoration

A method and apparatus comprising computer code configured to cause a processor or processors to receive an audio signal obtained from a microphone, input the audio signal into frequency-domain Kalman filter (FDKF), input the audio signal and an output from the FDKF into a neural network, estimate, based on the audio signal and the output from the FDKF, and removing feedback signals from the audio signal by the neural network, recover, by a codec receiving an output from the neural network, vocal quality of a target vocal signal, and output a version of the audio signal in which the target vocal signal is enhanced by removal of the feedback signals from the audio signal by the neural network and by recovery of the vocal quality by the codec.
Owner:TENCENT AMERICA LLC

Innovative karaoke sound system with combined neural network feedback suppression and coded vocal restoration

A method and apparatus comprising computer code configured to cause a processor or processors to receive an audio signal obtained from a microphone, input the audio signal into frequency-domain Kalman filter (FDKF), input the audio signal and an output from the FDKF into a neural network, estimate, based on the audio signal and the output from the FDKF, and removing feedback signals from the audio signal by the neural network, recover, by a codec receiving an output from the neural network, vocal quality of a target vocal signal, and output a version of the audio signal in which the target vocal signal is enhanced by removal of the feedback signals from the audio signal by the neural network and by recovery of the vocal quality by the codec.
Owner:TENCENT AMERICA LLC

AI-Based System for Vocal Lyric Modification

A system and method leveraging artificial intelligence to modify song lyrics while preserving the unique vocal characteristics of the original artist, including pitch, timbre, and expressive style. The system processes an input music file, separates vocal and instrumental components, and uses a trained singing voice model to synthesize a modified vocal track that seamlessly integrates new lyrics. The output retains the original artist's distinct vocal qualities, ensuring a realistic and natural rendition. Applications include content localization, personalized music production, and media streaming, enabling efficient lyric customization without requiring re-recording by the artist.
Owner:PARK JOHN CHOWHAN

An intelligent virtual robot

ActiveCN116901104BManipulatorVirtual robotTesting Methods
The application discloses a kind of intelligent virtual robots, including shell, lid is equipped on the upper end of shell, instruction input module, vocal music module and light module are equipped in the lid, wherein lid includes upper lid and lower lid, fixed plate is equipped between upper lid and lower lid, and vocal music module is set on the upper side of fixed plate, light module is set on the lower side of fixed plate, by such design can make structure more compact, and vocal music module is set on the upper side, can avoid the loss of tone quality;There is also a light-transmitting piece between the edge of upper lid and the edge of lower lid, and a light-transmitting groove is provided between the upper lid and the lower lid, which can guide the light emitted by the light module to the light-transmitting piece, so that the light of the light module can be displayed on the light-transmitting piece through the light-transmitting groove, thereby making the device have good light effect and vocal music effect, and the function is more complete.
Owner:SHENZHEN KIM DAI INTELLIGENCE INNOVATION TECHNOLOGY CO LTD

Artificial voice generation system

According to the present invention there is provided an artificial voice generation system designed to restore or augment speech for individuals experiencing voice loss or desiring an alternative vocal quality. The system comprises a controllable air pump, an airflow member delivering air to the user's oral cavity, and a sound generation member— such as a replaceable membrane cartridge—configured to vibrate and produce sound in response to airflow. A user interface allows manual or automatic modulation of airflow and voice parameters, enabling real-time adjustment of pitch, loudness, and voice quality, including male, female, or non-binary characteristics. The system may include sensors for pressure and airflow, anti-jamming features, and can be adapted for use with or without a neck stoma, making it suitable for a wide range of users, from laryngectomy patients to those with temporary or chronic voice loss. By providing a customisable, high-quality artificial voice source with improved naturalness and intelligibility, the invention addresses limitations of existing devices such as electrolarynxes and tracheoesophageal prostheses, offering a versatile and user-friendly solution for voice rehabilitation and enhancement.
Owner:LARONIX PTY LTD

Speech synthesis adaptation method and device, equipment and storage medium

The embodiment of the invention provides a speech synthesis adaptation method and device, equipment and a storage medium, which can be applied to customer service systems in the insurance fields of finance, medical treatment and the like. According to the method, basic training data and fine tuning data can be obtained firstly. And converting the training text into international phonetic symbols, and learning a cross-language general mapping relation on a phoneme level to obtain general mapping data. The base model is then trained based on the generic mapping data, and fine-tuned using the fine-tuning data. And performing online preference optimization on the basic model according to the prompt set and the multi-target reward function to obtain an adaptive model. According to the method, model fine tuning can be performed through multi-language IPA basic model training in combination with a small amount of paired data in a target language environment, and GRPO online preference optimization based on multi-index rewards is performed, so that the intelligibility, speaker consistency and tone quality of low-resource language synthesis are improved on the premise that large-scale parallel corpora are not needed.
Owner:PING AN TECH (SHENZHEN) CO LTD