Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

11 results about "Vocal quality" patented technology

Factors that determine quality. Production of original tone by the vocal folds (phonation); process of selection, reinforcement and damping of this tone by the resonators (throat, mouth, nose) Types of vocal quality. Breathy, strident, harsh, hypernasality, glottal fry, throaty, glottal attack, hoarse, hyponasality.

Method, device and equipment for voiceprint recognition system to resist attack and medium

PendingCN121963772AImprove migration abilityImprove stabilitySpeech analysisAttacker modelAlgorithm
The invention discloses a voiceprint recognition system attack resisting method and device, equipment and a medium, by constructing a shadow voice sample set and a substitution model, dependence on target speaker data and query capability is reduced, and privacy risk and query overhead are reduced; meanwhile, two types of substitution models are trained based on an information theory method, so that the mobility and stability of the models are improved, and the generalization ability is enhanced; besides, the attacker model adopts a specific architecture and combines an alternate training strategy, so that attack loss and prediction probability difference are optimized, imperceptibility and attack effectiveness are balanced, and tone quality reduction or attack failure is avoided; and finally, through systematic training and optimization, a more uniform evaluation caliber is expected to be provided, and more reliable guidance is provided for engineering landing and risk evaluation.
Owner:GUIZHOU UNIV

An adjustable frequency response microphone with integrated XLR and USB outputs

This invention discloses an adjustable frequency response microphone integrating XLR and USB outputs, comprising a microphone body, a dynamic microphone driver, a passive filter, a pass-through / preamp circuit, an XLR XLR interface, a USB module, and a function control circuit. The passive filter employs an LC circuit structure, supporting three frequency modes: low-frequency cutoff, mid-frequency boost, and flat response, effectively filtering out ambient noise and enhancing vocal quality. The pass-through / preamp circuit offers both pass-through and preamp modes, with adjustable preamp gain to accommodate dynamic microphone drivers of varying sensitivities, and can be externally or internally powered by phantom power. The USB module integrates a Type-C interface and an audio chip, supporting analog-to-digital / digital-to-analog conversion, real-time headphone monitoring, and external power supply. The function control circuit integrates an encoder and a touch panel, enabling adjustments such as volume, mixing, and mute, and enhances the interactive experience with an RGB lighting module. This invention combines the advantages of professional XLR output with convenient USB output, making it suitable for professional recording, live streaming, gaming, and other scenarios.
Owner:ZHAOQING HEJIA ELECTRONICS CO LTD

Innovative karaoke sound system with combined neural network feedback suppression and coded vocal restoration

A method and apparatus comprising computer code configured to cause a processor or processors to receive an audio signal obtained from a microphone, input the audio signal into frequency-domain Kalman filter (FDKF), input the audio signal and an output from the FDKF into a neural network, estimate, based on the audio signal and the output from the FDKF, and removing feedback signals from the audio signal by the neural network, recover, by a codec receiving an output from the neural network, vocal quality of a target vocal signal, and output a version of the audio signal in which the target vocal signal is enhanced by removal of the feedback signals from the audio signal by the neural network and by recovery of the vocal quality by the codec.
Owner:TENCENT AMERICA LLC

Innovative karaoke sound system with combined neural network feedback suppression and coded vocal restoration

A method and apparatus comprising computer code configured to cause a processor or processors to receive an audio signal obtained from a microphone, input the audio signal into frequency-domain Kalman filter (FDKF), input the audio signal and an output from the FDKF into a neural network, estimate, based on the audio signal and the output from the FDKF, and removing feedback signals from the audio signal by the neural network, recover, by a codec receiving an output from the neural network, vocal quality of a target vocal signal, and output a version of the audio signal in which the target vocal signal is enhanced by removal of the feedback signals from the audio signal by the neural network and by recovery of the vocal quality by the codec.
Owner:TENCENT AMERICA LLC

Voice providing device, voice providing method and program

To grasp a child's growing process from change of words and vocal quality which the child has uttered.SOLUTION: A CPU 21 of a terminal device 20: acquires voice data corresponding to adult language or a growth step term which satisfies an output condition from an utterance voice database 233, on the basis of a predetermined output condition; and outputs voice when the adult language or the growth step term which satisfies the output condition has been uttered by a predetermined speaker from a sound output unit 26 in time series, on the basis of the voice data.SELECTED DRAWING: Figure 7
Owner:CASIO COMPUTER CO LTD

AI-Based System for Vocal Lyric Modification

A system and method leveraging artificial intelligence to modify song lyrics while preserving the unique vocal characteristics of the original artist, including pitch, timbre, and expressive style. The system processes an input music file, separates vocal and instrumental components, and uses a trained singing voice model to synthesize a modified vocal track that seamlessly integrates new lyrics. The output retains the original artist's distinct vocal qualities, ensuring a realistic and natural rendition. Applications include content localization, personalized music production, and media streaming, enabling efficient lyric customization without requiring re-recording by the artist.
Owner:PARK JOHN CHOWHAN

An intelligent virtual robot

The application discloses a kind of intelligent virtual robots, including shell, lid is equipped on the upper end of shell, instruction input module, vocal music module and light module are equipped in the lid, wherein lid includes upper lid and lower lid, fixed plate is equipped between upper lid and lower lid, and vocal music module is set on the upper side of fixed plate, light module is set on the lower side of fixed plate, by such design can make structure more compact, and vocal music module is set on the upper side, can avoid the loss of tone quality;There is also a light-transmitting piece between the edge of upper lid and the edge of lower lid, and a light-transmitting groove is provided between the upper lid and the lower lid, which can guide the light emitted by the light module to the light-transmitting piece, so that the light of the light module can be displayed on the light-transmitting piece through the light-transmitting groove, thereby making the device have good light effect and vocal music effect, and the function is more complete.
Owner:SHENZHEN KIM DAI INTELLIGENCE INNOVATION TECHNOLOGY CO LTD

Artificial voice generation system

According to the present invention there is provided an artificial voice generation system designed to restore or augment speech for individuals experiencing voice loss or desiring an alternative vocal quality. The system comprises a controllable air pump, an airflow member delivering air to the user's oral cavity, and a sound generation member— such as a replaceable membrane cartridge—configured to vibrate and produce sound in response to airflow. A user interface allows manual or automatic modulation of airflow and voice parameters, enabling real-time adjustment of pitch, loudness, and voice quality, including male, female, or non-binary characteristics. The system may include sensors for pressure and airflow, anti-jamming features, and can be adapted for use with or without a neck stoma, making it suitable for a wide range of users, from laryngectomy patients to those with temporary or chronic voice loss. By providing a customisable, high-quality artificial voice source with improved naturalness and intelligibility, the invention addresses limitations of existing devices such as electrolarynxes and tracheoesophageal prostheses, offering a versatile and user-friendly solution for voice rehabilitation and enhancement.
Owner:LARONIX PTY LTD

Speech synthesis adaptation method and device, equipment and storage medium

The embodiment of the invention provides a speech synthesis adaptation method and device, equipment and a storage medium, which can be applied to customer service systems in the insurance fields of finance, medical treatment and the like. According to the method, basic training data and fine tuning data can be obtained firstly. And converting the training text into international phonetic symbols, and learning a cross-language general mapping relation on a phoneme level to obtain general mapping data. The base model is then trained based on the generic mapping data, and fine-tuned using the fine-tuning data. And performing online preference optimization on the basic model according to the prompt set and the multi-target reward function to obtain an adaptive model. According to the method, model fine tuning can be performed through multi-language IPA basic model training in combination with a small amount of paired data in a target language environment, and GRPO online preference optimization based on multi-index rewards is performed, so that the intelligibility, speaker consistency and tone quality of low-resource language synthesis are improved on the premise that large-scale parallel corpora are not needed.
Owner:PING AN TECH (SHENZHEN) CO LTD

Electric push rod with voice interaction function

The utility model provides an electric push rod with a voice interaction function, and relates to the technical field of electric push rods, the electric push rod comprises an assembly bottom shell and a sealed oil injection structure, one end of the top of the assembly bottom shell is fixedly provided with a motor, the other end of the top of the assembly bottom shell is fixedly connected with a telescopic sleeve, and the telescopic sleeve is fixedly connected with the motor through a connecting frame; a telescopic rod is arranged in the telescopic sleeve, a sealing oil injection structure is arranged at the top of the telescopic sleeve, a fixed top ring is fixedly connected to the top of the telescopic sleeve, an oil storage cavity is formed in the fixed top ring, an oil injection hole is formed in the top of the oil storage cavity, and a sealing oil plug is installed in the oil injection hole in a sealed mode. The electric push rod can prevent the problems of voice signal transmission distortion, tone quality reduction, circuit failure and the like caused by water or dust invasion, ensures that the electric push rod can still operate stably in a complex environment, can be self-lubricated, reduces the friction between the push rod and the shell, and enhances the smoothness of the push rod.
Owner:WUXI EAST YU XIANG SCI & TECH CO LTD

Dynamic calculation voice separation method based on human voice feature extraction and related equipment thereof

The invention provides a dynamic calculation voice separation method based on human voice feature extraction and related equipment thereof. The method comprises the following steps: acquiring voice data; performing feature extraction on the voice data, and generating a complexity score according to the extracted feature data; determining an exit point based on the complexity score, and executing voice separation processing at the exit point to obtain separated voice data and corresponding metadata; and performing tone quality optimization and semantic reasoning processing based on the separated voice data and the corresponding metadata to obtain target voice data. According to the method, the exit point of the voice separation network is dynamically determined through the combination of the complexity score and the equipment operation state, the computing resource consumption is reduced, differential optimization and semantic reasoning are carried out on the separation result by utilizing metadata, and the voice separation precision and the semantic understanding ability under the multi-speaker and complex environment are improved.
Owner:SHENZHEN STORYTELLING TECH CO LTD