Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

58 results about "Voice change" patented technology

A voice change or voice mutation, sometimes referred to as a voice break, commonly refers to the deepening of the voice of people as they reach puberty. Before puberty, both sexes have roughly similar vocal pitch, but during puberty the male voice typically deepens an octave, while the female voice usually deepens only by a few tones.

Method and system for changing voice outside vehicle

The invention discloses an external voice changing method and system, and the method comprises the steps: building a mapping relation between a voice changing scene and voice changing parameters, and enabling the voice changing scene to be used for representing an external environment where a vehicle is located; in response to the user operation, determining a target voice changing scene; obtaining a target voice changing parameter corresponding to the target voice changing scene based on the mapping relationship between the voice changing scene and the voice changing parameter; and performing voice changing processing on voice played by a loudspeaker outside the vehicle by adopting the target voice changing parameter. According to the method, the voice interaction mode of the vehicle and the external environment has scene perception and self-adaptive capabilities, and the voice outside the vehicle can be matched with the specific environment and is easier to accept.
Owner:CHONGQING HUAWEIXU ELECTRONICS CO LTD

Anti-fraud outbound call identification method for cross-number-segment voiceprint tracking

The invention relates to an anti-fraud outbound recognition method for cross-number-segment voiceprint tracking, and the method comprises the steps: introducing a time domain voiceprint adversarial extraction mechanism, and adding an adversarial voice discriminator, an active inhibition device, a channel, and the interference of non-identity factors of language emotion in voiceprint modeling; constructing a speaker constant feature residual channel: taking stable syllable fragments including vowels and gutto vowels in call content as a reference extraction window, and filtering emotion or speech speed driving components; outputting a steady-state voiceprint contour flow as a unique basis for subsequent cross-number identity fusion; designing a dynamic number fusion graph based on the voiceprint steady-state flow; each time of voiceprint appearance is regarded as a single time point node, and all similar historical voiceprints form a serial number merging path; a number time sequence edge weight function is set; the occurrence time interval, the use frequency and the regional jump of the new number and the old number are synthesized, and whether the number belongs to the same user or not is evaluated; and establishing atlas connection between the voiceprint main body node and all number nodes thereof.
Owner:SHIJIAZHUANG LINGYUE TECHNOLOGY CO LTD

Audio processing method and device, equipment and storage medium

The embodiment of the invention relates to an audio processing method and device, equipment and a storage medium. The method provided by the invention comprises the following steps: acquiring first media content input by a user, wherein the first media content comprises first audio content; and providing a second media content based on the user's selection of the target style, the second media content including a second audio content generated based on the first audio content, the second audio content having the same tone as the first audio content, and the second audio content having at least one audio attribute corresponding to the target style. Through the mode, the embodiment of the invention can improve the voice changing effect on the basis of keeping the tone quality.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

Audio communication method, audio conversion method, apparatus, electronic device, computer-readable storage medium, and computer program product

PCT designated stageWO2025237010A1Speech analysisComputer hardwareFeature coding
An audio communication method, an audio conversion method, a bitstream processing method, an apparatus, an electronic device, a computer-readable storage medium, and a computer program product. The audio communication method comprises: in response to a first communication request for an audio signal, acquiring, from a plurality of communication modes, a voice transformation mode for the audio signal (101); performing feature coding on the audio signal, so as to obtain a coded feature of the audio signal (102); acquiring a target timbre corresponding to the voice transformation mode, and determining a timbre feature of the target timbre (103); performing timbre conversion on the coded feature on the basis of the timbre feature, so as to obtain a target coded feature (104); and performing signal coding on the target coded feature, so as to obtain a target audio bitstream conforming to the target timbre, and transmitting the target audio bitstream to a decoding terminal (105).
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD

Audio processing method, computer equipment, storage medium and computer program product

The invention relates to an audio processing method, computer equipment, a storage medium and a computer program product. The method comprises the following steps: receiving audio data continuously acquired by audio acquisition equipment; under the condition that the current reasoning step is any non-first reasoning step, obtaining a to-be-processed fragment of the current reasoning step from the continuously collected audio data; inputting the to-be-processed segment of the current inference step into a pre-trained voice change model to obtain a model inference segment output by the voice change model; and according to the cached segments of the third duration, searching a segment connection point corresponding to the tail of the processed segment of the previous inference step in the model inference segment, and starting to intercept the segment of the second duration from the segment connection point in the model inference segment as the processed segment of the current inference step for playing. By adopting the method, the real-time voice changing effect can be improved.
Owner:TENCENT MUSIC ENTERTAINMENT TECH (SHENZHEN) CO LTD

Real-time voice inflexion method, terminal equipment and storage medium

The invention is suitable for the field of audio processing, and discloses a real-time voice inflexion method, terminal equipment and a storage medium. The real-time voice inflexion method comprises the steps of generating original voice data according to a real-time dialogue audio, and determining condition features and diversity features and filling data masks according to the original voice data; determining first tensor information according to the condition features, the diversity features and the filling data mask, and determining a speaker embedding vector according to the original voice data; determining second tensor information according to the first tensor information, the speaker embedding vector and the filling data mask; and generating a target timbre audio according to the second tensor information, the speaker embedding vector and the pitch frequency of the original voice data. According to the method, the reconstruction precision of the original timbre characteristics in the voice changing process is remarkably improved, the generated voice reaches the real person-like level in perception dimensions such as timbre similarity and intonation naturalness, and the voice changing authenticity of the real-time voice is improved.
Owner:SHENZHEN MAIFENG TECH CO LTD

Emotion and behavior anomaly detection and early warning method and device, equipment and storage medium

The invention discloses an emotion and behavior anomaly detection and early warning method and device, equipment and a storage medium, and the method employs a wearable detection device which can detect the speech and behavior of a server, the speech, expression and posture of a serviced person, and the environment where the serviced person is located. And the emotion and the holding behavior of the server and the serviced person are comprehensively analyzed, so that early warning is carried out on the severe emotion change and the violent behavior. Through a camera and a microphone on the multi-mode cooperative emotion sensing module, emotion changes and voice changes of a server and a serviced person can be clearly recorded, and the emotions and actions of the server and the serviced person are analyzed through the multi-mode cooperative emotion sensing module. On one hand, emotion and behavior changes of a server and a serviced person are recorded to provide data support for subsequent better services, and on the other hand, events can be better restored through the recorded data after accidents occur. The invention relates to a character and environment detection technology.
Owner:SOUTH CHINA UNIV OF TECH +1

Voice interaction method and system of live broadcast room, storage medium and computer equipment

PendingCN121771419ASelective content distributionAudience interactionEngineering
The invention provides a voice interaction method and system for a live broadcast room, a storage medium and computer equipment, and the method comprises the steps: obtaining voice-changing audio data and interaction answer information corresponding to the voice-changing audio data in response to an instruction that an anchor client starts a voice-changing playing method through a live broadcast server, sending the variable sound audio data to audience clients in a live broadcasting room; the audience client plays the variable sound audio data, obtains audience interaction answer information in response to an answer selection instruction, and sends the audience interaction answer information to the live broadcast server; and the live broadcast server generates interaction result data according to the interaction answer information and the audience interaction answer information. Therefore, by applying the technical scheme of the invention, the interestingness and the interaction efficiency of live broadcast can be effectively improved, and the audiovisual experience of audiences in the live broadcast room and the user stickiness are enhanced.
Owner:GUANGZHOU FANGGUI INFORMATION TECHNOLOGY CO LTD

Voice changer (with amplifier)

ActiveCN309456045SVoice changeAmplifier
1. The name of this design product: Voice changer (with amplifier). 2. Purpose of this design product: used to change sound. 3. The key design point of this design product lies in the combination of shape and pattern. 4. The picture or photo that best illustrates the design points: three-dimensional picture.
Owner:DONGGUAN KANGBEISHI INTELLIGENT TECH CO LTD

Toy (recording sound changing amplifying singing machine)

ActiveCN309623907SEngineeringVoice change
1. The name of the design product: toy (recording voice amplifying singing machine). 2. The use of the design product: toy for entertainment and play. 3. The design points of the design product: in shape. 4. The picture or photo that best indicates the design points: assembly 1 perspective view.
Owner:林镇强

Variable voice authentic identification and traceability method and system based on multi-task learning

The invention discloses a variable voice authentic identification and traceability method and system based on multi-task learning, and relates to the field of artificial intelligence and computer security. Obtaining an original voice sample; constructing a variable voice authentic identification and traceability model; inputting the original voice sample into a final variable voice authentic identification and traceability model to obtain a voice forgery probability and an original speaker; the variable voice authentic identification and traceability model integrates a feature extraction module, a shared encoder module, an authentic identification task branch module and a traceability task branch module; according to the invention, a joint learning framework capable of collaborative optimization and mutual promotion is constructed, multi-task learning is utilized to improve the authenticity identification accuracy and traceability of the voice-changing voice at the same time, and on the basis of high-precision discrimination of whether an original voice sample is forged or not, the original speaker voiceprint in the voice-changing voice is further reversely recovered, so that the voice-changing voice identification accuracy and traceability of the voice-changing voice are improved. And a high-performance defense technical scheme for variable voices is realized.
Owner:ZHEJIANG UNIV

Personalized electrical stimulation treatment scheme generation system and method

The invention belongs to the field of auxiliary medical treatment, and provides a personalized electrical stimulation treatment scheme generation system and method, and the method comprises the steps: obtaining a voice signal of a patient, and extracting voice dynamic change information; acquiring neuromyoelectric signals, heart rate signals and motion information of the patient, and extracting physiological change information; performing association modeling on the voice change information and the physiological change information to generate fusion features for representing a multi-part tremor mode of the patient; performing dynamic generation of personalized electrical stimulation parameters based on the fused features; constraining the dynamic change of the current amplitude, the stimulation frequency and the pulse width within an individually set safety threshold range, and performing protective adjustment in combination with real-time feedback of a patient; and the electrical stimulation module is used for implementing electrical stimulation on a target nerve part in the radial nerve, the median nerve or the ulnar nerve according to the parameters issued by the decision module.
Owner:FANSKY CO LTD

Voice changing method, apparatus and device for speech dialog in live-streaming room, and medium

PCT designated stageWO2026138260A1Voice changeSpoken dialog
A voice changing method, apparatus and device for a speech dialog in a live-streaming room, and a medium, which relate to the field of network live streaming. The method comprises: in response to a speech speaking event of a speaker user in a live-streaming room, detecting and determining a speech segment in target audio data; on the basis of a preset target pitch value, performing tone-change processing on a segment pitch feature of the speech segment, so as to obtain an optimized pitch feature; on the basis of the optimized pitch feature and a target timbre feature, performing voice-change processing on the speech segment, so as to obtain a voice-changed segment; and replacing a corresponding speech segment in the target audio data with the voice-changed segment, so as to obtain voice-changed audio data, and sending the voice-changed audio data to a receiver user in the live-streaming room. The present application significantly improves the performance of voice changing technology, and solves the problems in the conventional technology in the aspects of real-time performance, personalized services, pitch adjustment naturalness, real timbre restoration, etc.
Owner:GUANGZHOU FANGGUI INFORMATION TECHNOLOGY CO LTD

A voice-changing multifunctional alarm

The present invention discloses a sound-changing multifunctional alarm, which relates to the technical field of alarms. The present invention comprises an outer ring wheel, an alarm is installed on the inner side of the outer ring wheel, an outer spoiler is provided in an annular shape inside the outer ring wheel, an inner ring wheel is rotatably installed inside the outer ring wheel, an inner spoiler is provided inside the inner ring wheel, and a fan is rotatably installed inside the outer ring wheel. In the present invention, air is taken into the inner and outer ring wheels by the rotation of the fan, and discharged through the inner and outer spoiler blocks. The airflow is broken up by the inner and outer spoiler blocks and converted into audible sound at the correct frequency. At the same time, the airflow is blocked by the inner and outer spoiler blocks, resulting in a larger pressure difference between the inner and outer ring wheels, thereby making the sound louder, thereby achieving the effect of alarming. When used in conjunction with the alarm, different alarm sounds can be generated, and the alarm can be applied to different environments.
Owner:SANXING QILONG SCI & TECH IND JIANGXI

Multifunctional sound device

Disclosed in the present invention is a multifunctional sound device, comprising a silencing housing and a microphone, wherein a bracket is connected to the lower part of the silencing housing, an air chamber is provided inside the silencing housing, and mini air-passing tubing is embedded in the silencing housing; an isolation box is connected around the microphone; the microphone is connected to a data cable which is embedded in the silencing housing; an independent data switch is further provided outside the data cable; the microphone is further provided with a receiver that can be wirelessly connected; and a sound sensor is connected beside the microphone. The technical solution of the present invention can realize confidential speaking, noise-reducing speaking and voice-changing speaking.
Owner:YANG HUA

Apparatus for diagnosing disease causing voice and swallowing disorders and method for diagnosing same

An apparatus for diagnosing a disease and a method for diagnosing a disease, in which: a plurality of voice signals are received to generate a first image signal and a second image signal which are image signals for each voice signal; a plurality of disease probability information for a target disease causing a voice change are extracted by using an artificial intelligence model determined according to the type of each voice signal and a generation method used to generate each image signal for the first image signal and the second image signal for each voice signal; and it is determined whether the target disease is negative or positive on the basis of the plurality of disease probability information.
Owner:THE CATHOLIC UNIV OF KOREA IND ACADEMIC COOP FOUND +1

Karaoke program, karaoke equipment, and karaoke scoring method

The purpose is to appropriately evaluate voice-changed audio. [Solution] The system is characterized by performing a playback process that plays music based on music information, a first input process that receives voice-changed audio, a second input process that receives audio that is not voice-changed, a sound output process that outputs the played music and voice-changed audio, a scoring process that scores at least a portion of a plurality of scoring items, namely a first scoring item, using the voice-changed audio, and a result output process that outputs the results of the scoring process.
Owner:BROTHER KOGYO KK

Karaoke program and karaoke device

To allow a user to hear voice that has been voice change-processed and music sound to be reproduced without discomfort.SOLUTION: A Karaoke program executable by an information processing device performs reproduction processing for reproducing music sound based on music information, first input processing for inputting voice that has been voice change-processed, first output processing for outputting the reproduced music sound to a singer, and second output processing for pitch-correcting and outputting the reproduced music sound based on a pitch difference between the voice that has been voice change-processed and the voice before being voice change-processed.SELECTED DRAWING: Figure 6
Owner:BROTHER KOGYO KK

Southern Fujian Chinese opera voice changing method and system based on RVC network

The invention provides a southern Fujian Chinese opera voice changing method and system based on an RVC network in the technical field of artificial intelligence and voice crossing. The method comprises the steps that S1, a large number of historical Chinese audios and historical southern Fujian Chinese opera audios are acquired to construct a data set; s2, creating a semantic dictionary used for storing the corresponding relationship among the southern Fujian pinyin, the Chinese pinyin and the digital sequence, and the corresponding relationship between the pitch and the digital sequence; the corresponding relations are separated through preset section separators; s3, creating a southern Fujian opera sound change model, and setting a loss function of the southern Fujian opera sound change model; s4, training a southern Fujian Chinese opera voice change model through the data set and the loss function; and step S5, deploying the trained Southern Fujian Chinese opera voice changing model, and carrying out Southern Fujian Chinese opera voice changing operation through the deployed Southern Fujian Chinese opera voice changing model. The method has the advantage that the sound change quality of the Chinese opera of the southern Fujian is greatly improved.
Owner:XIAMEN UNIV

Method, apparatus, device, and storage medium for audio processing

Embodiments of the disclosure relate to a method, apparatus, device, and storage medium for audio processing. The method provided herein includes: obtaining a first media content input by a user, the first media content including a first audio content corresponding to a singing content; and providing a second media content based on a selection of a target timbre by the user, the second media content including a second audio content corresponding to the singing content, and the second audio content corresponding to the selected target timbre. In this way, the embodiments of the disclosure can convert the first audio content in the audio corresponding to the singing content into a specified timbre, thereby improving the voice changing effect while retaining the timbre.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD +1

Toy (dinosaur voice-changing speaker)

ActiveCN309513601SEngineeringVoice change
1. Name of the product under this design: Toy (Dinosaur Voice Changing Loudspeaker). 2. Purpose of the product of this design: a toy used for entertainment and play. 3. The key point of the design of this product lies in its shape. 4. The picture or photo that best illustrates the design points: Stereoscopic drawing 1.
Owner:林镇强

A Fraud Prevention Communication Method Based on Voice Changer Recognition

This invention relates to the field of voiceprint recognition technology, and more particularly to an anti-fraud communication method based on voice-changing recognition. The method includes the following steps: acquiring multi-dimensional voice feature data of the user, including timbre, frequency features, speech rhythm, emotional features, and pause patterns; predicting the probability of voice changing based on the multi-dimensional voice feature data to obtain voiceprint feature prediction data; training a deep learning model based on the voiceprint feature prediction data, and mapping the voiceprint state and features under different call conditions to obtain voiceprint feature map data; and simulating and analyzing voice-changing recognition behavior under different call conditions based on the voiceprint feature map data to obtain voice-changing recognition simulation data. This invention and method effectively reduce the impact of the environment on recognition by acquiring real-time call environment noise data and network quality data, constructing a call quality model, and performing cross-scenario optimization and real-time adjustment.
Owner:MINAMI ACOUSTICS LTD

System

To provide a system for supporting the intellectual training and growth of a child while reducing the labor of a parent.SOLUTION: The voice recognition system includes a means for acquiring voice uttered by a user, a voice recognition means for converting the acquired voice into text data, and a communication means for transmitting the converted text data. A system comprising: language analysis means for extracting a user's intention and emotional state; response generation means for generating an appropriate response based on the extracted intention and emotional state; communication means for receiving the generated response as text data; speech synthesis means for converting the received response text into voice data; and voice changer means for applying character voices to the converted voice data.SELECTED DRAWING: Figure 1
Owner:SOFTBANK GROUP CORP

A method and system for comparing Chinese oral pronunciation based on speech recognition

PendingCN122454977ASpoken languageRhythm pattern
The application provides a Chinese oral pronunciation comparison method and system based on voice recognition, relates to the field of voice recognition, obtains text information to be pronounced and performs linguistic rule analysis, simulates the continuous motion track of the pronunciation organ and the acoustic adjustment law, thereby generating a standard reference voice containing voice continuity and rhythm patterns, compares the acoustic characteristics of the learner's pronunciation with the acoustic characteristics of the standard reference voice, and determines whether the difference is a pronunciation defect according to the natural voice changes reflected in the standard reference voice; the application can effectively distinguish the subtle differences in the learner's pronunciation that belong to natural voice changes from the real pronunciation defects, solve the technical problem that the existing system misjudges the natural coarticulation effect as a pronunciation error, improve the authenticity and authenticity of oral expression, overcome the deficiency that the system feedback is disconnected with the actual perception in the prior art, and significantly improve the efficiency and effect of Chinese oral learning.
Owner:HUBEI UNIV OF EDUCATION

Real-time voice changing method, electronic equipment and computer program product

The invention discloses a real-time voice changing method, electronic equipment and a computer program product, and relates to the technical field of artificial intelligence. The method comprises the following steps of: respectively extracting audio content characteristics, fundamental frequency characteristics and audio energy characteristics from source audio, and performing segmented accumulation phase operation on the fundamental frequency characteristics in a time dimension to generate waveform signal data subjected to streaming processing; a causal neural network layer serving as a neural vocoder is utilized to carry out dimension reduction alignment on audio content features, waveform signal data, audio energy features and target tone information under time sequence constraint, and a time domain waveform signal is reconstructed through scale expansion operation and multi-path feature transformation operation. And the target audio data is converted into the target tone as the source audio. According to the invention, the problem that low computing resource occupation, high sound quality and real-time sound changing cannot be considered in the prior art can be solved, and high-sound-quality real-time sound changing can be realized under the condition of low computing resource occupation.
Owner:CHENGDU YIWO TECH DEV CO LTD

Audio processing method and apparatus, computer device, readable storage medium, and program product

PCT designated stageWO2025251757A1Speech analysisTransmissionEngineeringVoice change
An audio processing method, executed by a computer device. The method comprises: collecting an audio signal in real time in the process of performing audio interaction with an audio interaction terminal by using an identity of an object identifier (502); calculating a real-time voiceprint vector of the audio signal collected in real time, and comparing the calculated real-time voiceprint vector with a reference voiceprint vector registered for the object identifier and belonging to a registered object to obtain a comparison result (504); when it is determined, according to the comparison result, that current speaking voice of the audio signal is from the registered object, sending the audio signal to the audio interaction terminal (506); and when it is determined, according to the comparison result, that the current speaking voice of the audio signal is not from the registered object, performing voice alteration processing on the audio signal to obtain a voice-altered audio signal, and sending the voice-altered audio signal to the audio interaction terminal (508).
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD

Harmony processing method, device, equipment and medium

Embodiments of the present disclosure relate to a harmony processing method, apparatus, device, and medium. The method includes: in response to a triggering operation on a target harmony control, obtaining a harmony interval corresponding to the target harmony control; performing pitch shifting processing on a first sound of the original input according to the harmony interval to obtain a second sound; wherein the interval between the first sound and the second sound is the harmony interval; generating a target audio based on the first sound and the second sound, and the first sound and the second sound are presented as different harmony parts in the target audio. Thus, based on the triggering of the harmony control, different harmony sounds are realized according to the corresponding harmony parameters, which improves the human-computer interaction feeling, reduces the harmony addition cost, enriches the diversity of sound playback, improves the interestingness of sound playback, and adds a harmony effect based on the original first sound played, which improves the smoothness and naturalness of harmony addition.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

Throat lesion monitoring system based on voice recognition

InactiveCN120048291ASpeech analysisSensorsDisease monitoringAbnormal voice
The invention relates to the technical field of sound analysis, in particular to a throat lesion monitoring system based on sound recognition, which comprises a sound quality optimization module, a sound feature analysis module, an abnormality evaluation module and a treatment effect evaluation module. According to the invention, by utilizing the historical sound data of the user and comparing and evaluating the deviation of the current sound features, the monitoring accuracy is improved, the detection of early lesions is more reliable, the deviation detection threshold is adjusted according to the age and gender differences of individuals, the sound anomaly detection is more accurate, the sound anomaly of the user can be found in time, and the user experience is improved. The system has the advantages that the system is simple in structure and convenient to use, false alarm and missing alarm are reduced, real-time and visual treatment feedback is provided by collecting sound data before and after treatment, analyzing the sound change degree, combining the treatment time and evaluating the treatment effect, data support is provided for doctors, measure adjustment in the treatment process is optimized, and the disease monitoring and management fineness and effect are improved.
Owner:THE FIRST PEOPLES HOSPITAL OF NANTONG

Intelligent linkage bid evaluation base management method and system

This invention provides an intelligent, interconnected management method and system for bid evaluation bases. By deeply integrating multiple advanced technologies such as facial recognition, ID card recognition, voice changing, location monitoring, and big data analysis, it forms an organic whole. This achieves comprehensive, intelligent management of the entire bid evaluation base process, from pre-bid preparation, personnel entry, review, to personnel departure. Furthermore, it utilizes data analysis and intelligent algorithms for full-process monitoring and early warning of the bid evaluation process, enabling real-time and precise supervision. It can promptly detect and warn of various abnormal behaviors and potential risks, providing strong decision-making support for regulatory departments. This changes the previous situation of lagging and passive supervision, greatly improving regulatory efficiency and effectiveness.
Owner:BEIJING JINGNENG TENDERING & COLLECTIVE PROCUREMENT CENT CO LTD

Toy (voice-changing microphone)

ActiveCN310014778SVoice changeAudiology
1. Name of the product in this design: Toy (Voice Changer Microphone). 2. Purpose of this design: for children to play with. 3. The key design feature of this product is its shape. 4. The image or photograph that best illustrates the key design points: Design 1 3D view. 5. Design 1 is designated as the basic design.
Owner:兰建东