Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

39 results about "Sound separation" patented technology

Separation of Sound Sources. A fundamental problem in sound separation is that when two or more sounds overlap with each other in time and frequency, separation is diffucult and there is no general method to resolve the component sounds.

Power equipment fault diagnosis prediction method and device and electronic equipment

The invention relates to a power equipment fault diagnosis prediction method and device and electronic equipment, and relates to the technical field of artificial intelligence, and the method comprises the steps: collecting sound data and electrical data of a plurality of pieces of power equipment, positioning noise data from the sound data, and extracting fault features according to a time-varying characteristic curve of the sound data of the positioned noise data; positioning the fault time nodes with the extracted fault features, and matching the electrical data corresponding to the fault time nodes to obtain fault electrical data; separating noise data in the sound data by using a sound separation model to obtain the noise data of each power device, and determining the fault of each power device according to the noise data of each power device, the time-varying characteristic curve and the matching degree of the fault electrical data; and predicting a fault trend according to the data exception trend of the fault type of the fault. By applying the scheme of the invention, the accuracy and reliability of partial discharge detection can be improved, and the risk of missing detection or misjudgment can be reduced.
Owner:CHANGSHU INSTITUTE OF TECHNOLOGY

Sound Separation Based on Distance Estimation using Machine Learning Models

A computer-implemented method of applying a trained neural network for sound separation based on distance estimation is provided. The method includes receiving, by an audio input component of a computing device, an audio mixture from one or more sources. The method includes predicting, by a trained distance estimation neural network and based on the audio mixture, respective distances of the one or more sources from the audio input component. The method includes determining one or more near sounds and one or more far sounds based on the respective distances. The near sounds correspond to sources that are located within a threshold distance of the audio input component, and the far sounds correspond to sources that are not located within the threshold distance of the audio input component. The method includes providing the predicted one or more near sounds.
Owner:GOOGLE LLC

Method and system for controlling operation of air conditioning equipment and storage medium

The invention relates to the field of intelligent environment control, in particular to a method and system for controlling operation of air conditioning equipment and a storage medium. The method comprises the following steps: acquiring external sound; based on a voiceprint library of the air conditioning equipment, air conditioning equipment noise and environment noise are separated from the external sound in a frequency domain; setting, for each frequency in the spectrum of the external sound, a weight indicating a corresponding tolerance of noise at that frequency; on the basis of the set weight, whether the noise of the air conditioning equipment at the frequency affects the user experience or not is determined for each frequency; and when the noise of the air conditioning equipment at any frequency affects the user experience, operation of the air conditioning equipment is adjusted.
Owner:HONEYWELL ENVIRONMENTAL & COMBUSTION CONTROLS (TIANJIN) CO LTD

Sound separation and target sound extraction method based on unified architecture

The invention discloses a sound separation and target sound extraction method based on a unified architecture, and relates to the technical field of audio signal processing, and the method comprises the steps: in a first training stage, employing an attractor network to estimate the number of sound sources in a mixed audio signal, and generating attractor embedding; the attractors are embedded and input into the separation backbone network, and separated sound is generated; in the second training stage, the multi-mode clue processing network is adopted to process clue information of multiple modes, and clue embedding is generated; randomly selecting attractor embedding or clue embedding as input for separating the backbone network; in the training stage, an alignment loss function between attractor embedding and clue embedding is calculated; in the reasoning stage, if there is no clue information, a sound separation task is executed; and if one to three pieces of clue information exist, executing a target sound extraction task. The task is flexibly selected and executed according to the input clue condition, better adaptability and flexibility are achieved, and the method is suitable for complex sound scenes.
Owner:SHANGHAI JIAOTONG UNIV

Audio signal processing device

The present invention addresses the problem of providing an "audio signal processing device" that achieves a good sense of localization with respect to a sound having a narrow frequency band. [Solution] The present invention is provided with signal processing units provided so as to correspond to respective channels of audio signals of a plurality of channels, and each signal processing unit has: a target sound separation unit (411) that separates a target sound signal indicating a target sound, which is a sound to be subjected to localization enhancement, from a source sound signal, which is an audio signal of the corresponding channel; a sense-of-localization-enhanced sound generation unit that generates a sense-of-localization-enhanced sound signal in which the frequency of the target sound signal separated by the target sound separation unit (411) is multiplied by k (k is an integer of 2 or more); and a synthesis unit that synthesizes and outputs the localization-enhanced sound signal generated by the localization-enhanced sound generation unit and the source sound signal.
Owner:ALPS ALPINE CO LTD

A sound separation method and system based on vector microphones

The present application relates to the field of audio processing, in particular to a sound separation method and system based on a vector microphone, the method comprising: constructing a fourth-order cumulant matrix of a sound signal received by the vector microphone, solving sound source direction angle estimation and sound source pitch angle estimation of each incident signal; constructing a spatial characteristic matrix of each incident signal at different frequencies; based on the spatial characteristic matrix, calculating a spatial correlation matrix estimation of the sound signal at different frequencies and different frame numbers, and constructing a spatial correlation matrix of the sound signal at different frequencies and different frame numbers; constructing an orthogonal MNMF model; adjusting the variables of the MNMF through the multiplication update method, so that the difference between the spatial correlation matrix estimation and the spatial correlation matrix is less than a preset target, to obtain an optimal spatial correlation matrix estimation; and separating the sound signal based on the optimal spatial correlation matrix. The present application effectively solves the sound separation problem when the distance between sound sources is close, and improves the accuracy of sound separation.
Owner:ANHUI UNIV +1

Individual perception multi-bird sound separation method and system based on identity embedding

The invention belongs to the field of sound separation processing, and discloses an individual perception multi-bird sound separation method and system based on identity embedding, and the method comprises the steps: S1, carrying out the preprocessing of an audio segment containing bird sound, obtaining a preprocessed audio segment, and obtaining a time domain waveform corresponding to the audio segment; s2, carrying out separation processing on the time domain waveform to obtain a separated individual audio waveform; and S3, based on the separated individual audio waveform, carrying out classification identification on the bird sound. According to the method, under the conditions that voiceprint differences of individuals of the same type of birds are extremely small and sound sources are highly overlapped, stable individual consistency and separation precision can still be kept. According to the method, the technical bottleneck of identity drift and pseudo separation of a traditional model in long-time sequence monitoring is solved, refinement and automation of individual-level ecological acoustic analysis are also realized, and a feasible technical path is provided for field long-term ecological monitoring and acoustic big data processing.
Owner:GUANGZHOU UNIVERSITY

Three-dimensional point cloud fused sound source separation method and device for power transformation main equipment

The invention relates to a three-dimensional point cloud fused power transformation main equipment sound source separation method, which comprises the following steps of: acquiring high-density point cloud data, preprocessing the point cloud data, and performing equipment segmentation; extracting sound signal features; fusion features are obtained; inputting the fusion features into a sound source separation network based on Transform, and carrying out sound signal separation to obtain separated sound signals; and inputting the separated sound signals into a sound source separation network based on Transform to carry out iteration for multiple times to optimize a separation result. The invention provides a novel multi-sound-source separation method, and the method can remarkably improve the sound source positioning and recognition precision through the combination of the three-dimensional point cloud and sound signals. Compared with a traditional sound separation method, the method has the advantages that a sound source can be positioned more accurately by means of the three-dimensional point cloud data, noise interference is reduced, the reliability of substation equipment fault diagnosis is improved, and the method has wide application prospects and particularly has important significance in the aspects of automatic monitoring and maintenance of a power system.
Owner:ANHUI NANRUI JIYUAN POWER GRID TECH CO LTD

Mechanical arm deception attack side channel detection method and system based on acoustic features

The application discloses a mechanical arm deception attack side channel detection method based on acoustic characteristics and belongs to the technical field of deep learning, which comprises the following steps: constructing a training recognition model; receiving the collected target audio, performing a sound processing step operation to extract features, performing a training recognition step operation to output predicted motion data; comparing the predicted motion data with the obtained corresponding real-time motion data, and evaluating whether the difference between each pair of parameters exceeds a preset threshold value. The application uses sound separation technology to separate mixed multi-axis sound into different individual axis sound, constructs an acoustic motion information recognition model of each motion joint through a collaborative physical and data-driven method, identifies the target sound according to the recognition model, obtains the predicted motion information of each axis, compares the predicted motion information with the collected real-time motion command, combines a threshold value to determine whether a deception attack is received, and realizes mechanical arm deception attack side channel detection.
Owner:GUANGXI UNIV

NOISE SOURCE SEPARATION USING ANGLE POSITION

Systems and methods for audio source separation. A deep learning-based system uses azimuth angle position to separate an audio signal originating from a selected location from other sounds. Techniques for directing a virtual microphone towards a selected speaker are revealed. A deep learning-based audio regression method, which can be implemented as a neural network, learns to separate different speakers by effectively utilizing the spectral and spatial properties of all sources. The neural network can focus on multiple sources in multiple respective target directions and filter out other sounds. A user can select which source to hear. The network can use the time-domain signal and a frequency-domain signal to separate the target signal and produce separate audio outputs.The direction of the selected speaker relative to the microphone array can be entered into the system as a vector.
Owner:INTEL CORP

Deep learning-based sound isolation method, device and storage medium

The application discloses a sound isolation method and device based on deep learning and a storage medium. The method comprises the following steps: obtaining an audio file used for constructing a DeepAudioSep model and preprocessing the audio file used for constructing the DeepAudioSep model; constructing the DeepAudioSep model and training the DeepAudioSep model, wherein the DeepAudioSep model comprises one mixed source input and ten isolated source outputs; and performing sound separation through the DeepAudioSep model. The application introduces the data driving and deep learning idea into sound separation and noise isolation processing, improves the sound separation and noise isolation processing capability in the environmental monitoring field, and therefore has wide noise processing prospects and practical value.
Owner:ZHUHAI GAOLING INFORMATION TECH COLTD

System and method for automatic detection of disease-associated respiratory sounds

A method to detect and analyze cough parameters remotely and in the background through the user's device in everyday life. The method involves consecutive or selective execution of 4 stages, which solve a variety of the following tasks: Detection of coughing events in sound, in environmental noise. Separation of the sound containing the cough into the sound of the cough and the rest of the sounds, even when extraneous sounds occurred during the cough. Identify the user by the sound of the cough to prevent analyzing the sounds of the cough that do not belong to the user (for example when another person coughs nearby). Assessment of cough characteristics (wet / dry, severity, duration, number of spasms, etc.). The method can work on wearable devices, smart devices, personal computers, laptops in 24 / 7 mode or in a selected range of time, and has a high energy efficiency.
Owner:INSUBIQ INC

Audio signal processing unit

We provide an "audio signal processing device" that achieves good localization for sounds with a narrow frequency band. [Solution] The system has a signal processing unit corresponding to each channel of a multi-channel audio signal. Each signal processing unit includes a target sound separation unit 411 that separates a target sound signal, which is the sound to be targeted for localization enhancement, from a source sound signal, which is the audio signal of the corresponding channel; a localization enhancement sound generation unit that generates a localization enhancement sound signal by multiplying the frequency of the target sound signal separated by the target sound separation unit 411 by k (where k is an integer of 2 or more); and a synthesis unit that synthesizes the localization enhancement sound signal generated by the localization enhancement sound generation unit and the source sound signal and outputs the result.
Owner:ALPS ALPINE CO LTD

Adaptive frequency band optimization chorus sound separation method and system based on harmonic preservation

The present application relates to a kind of adaptive band optimization chorus sound separation method and system based on harmonic maintenance, first, the signal is collected by microphone array and time delay compensation and multi-channel time-frequency transformation are carried out;Sound source spatial positioning and initial band division are carried out using MUSIC algorithm;Using the greedy optimization strategy, the band boundary is dynamically adjusted and optimized by maximizing the objective function including signal-to-noise ratio and harmonic integrity measure;Then, the fundamental frequency is obtained by harmonic analysis and detection, and the time-varying harmonic trajectory is tracked across frames by Viterbi algorithm;On this basis, combined with MVDR beam forming and the introduction of harmonic maintenance factor, enhanced time-frequency representation is obtained by soft time-frequency mask;Finally, inverse short-time Fourier transform time-domain reconstruction and synchronous harmonic phase are carried out.The present application realizes the accurate allocation of voice part spectrum, effectively reduces the overlap interference, protects the overtone column of human voice while accurately separating, significantly improves the timbre fidelity and naturalness of separated voice part.
Owner:SHANGHAI UNIVERSITY OF ELECTRIC POWER

Air conduction hearing aid coincidence sound separation method

The invention relates to the field of air conduction hearing aids, in particular to an air conduction hearing aid coincidence sound separation method, which comprises a hearing aid module, a coincidence sound processing module, a telescopic module, a sound insulation module and an auxiliary hearing aid module, and is characterized in that the hearing aid module comprises an external ear fixing assembly, a control assembly and a broadcaster assembly; the coincident sound processing module comprises a processing chip, a coincident sound separating module and a coincident sound playing module; the telescopic module comprises a storage assembly and a telescopic assembly, the sound transmitter assembly is controlled to extend into the ear canal, the storage assembly is opened, the inflation assembly is controlled to expand the inflation air bag to isolate external sound, the bone conduction assembly is driven to abut against the inner ear canal, and then coincident sound is separated through the processing chip; and then the positive sound part is played through the separated broadcaster assembly, and the bone conduction assembly plays the coincident sound part, so that the situation that the same player plays too disorderly is avoided, and then the situation that information in the coincident sound is missed under the normal use condition can be avoided.
Owner:LEFT POINT HEALTH IND (SHENZHEN) CO LTD

Sound directional frequency conversion method and system, hearing aid equipment, medium and hearing aid robot

The invention provides a sound directional frequency conversion method and system, hearing aid equipment, a medium and a hearing aid robot, and the method comprises the steps: carrying out the face recognition of the image information of a person, and obtaining a corresponding face recognition result; performing beam forming on the microphone array sound, and performing pronunciation object recognition and pronunciation object direction recognition according to a beam forming result, or performing scene recognition on external input sound; and performing sound separation based on the sound recognition result, matching a current frequency conversion scheme from a frequency conversion scheme database stored in the cloud according to the face recognition result and the sound separation result, performing directional frequency conversion according to the current frequency conversion scheme by using a sound frequency conversion algorithm stored in the local end, and outputting directional frequency conversion sound to the target person. Through directional frequency conversion of sound, high-quality auditory experience can be provided for hearing impairment patients, the hearing impairment patients can be better integrated into the society and life, and the system has wide application prospects in the fields of old-age care, medical treatment and the like.
Owner:HAINAN QINGWEN YAYIN TECHNOLOGY CO LTD

Passive respiratory disease early warning equipment based on respiratory audio spectrum feature separation and bidirectional long short-term memory network

The invention discloses a passive respiratory disease early warning device based on respiratory audio spectrum feature separation and a bidirectional long short-term memory network, which is characterized in that signal preprocessing is triggered when a posture is static by continuously collecting respiratory sound and synchronizing body position and motion state data; cardiopulmonary sounds are separated in real time by adopting an online non-negative matrix factorization technology, pure respiratory signals are extracted, and a dynamic signal environment is adapted through an incremental updating mechanism; key frequency band features are extracted in combination with wavelet transform time-frequency analysis, and breathing micro-variation features such as expiratory phase extension and wet rale are quantized by using a dynamic time bending algorithm; a two-way LSTM fused with an external attention mechanism is introduced to model breathing time sequence characteristics, key frequency domain information is focused, abnormal probabilities of pneumonia, asthma and chronic obstructive pulmonary disease are output, and when a preset early warning condition is met, a prompt is triggered; respiration feature extraction and desensitization are completed through local edge calculation, and only abnormal fragment ciphertexts are uploaded to guarantee data privacy. The method solves the problems that in the prior art, breathing sound separation precision is insufficient, single-mode analysis is limited, time sequence modeling adaptability is poor and the like, and zero-intervention and high-precision early passive monitoring of respiratory system diseases is achieved.
Owner:FUSHOUKANG (SHANGHAI) FAMILY SERVICES CO LTD

Multi-target sleep monitoring method and device based on snoring sound separation

This invention discloses a multi-target sleep monitoring method and device based on snoring separation. The method includes: constructing independent respiratory wave time-stamped sequences corresponding to each monitored object based on the radio frequency signal corresponding to the target monitoring area; constructing a mixed snoring time-stamped sequence corresponding to all monitored objects based on the mixed environmental audio signal corresponding to the target monitoring area; performing cross-modal correlation matching between the mixed snoring time-stamped sequence and each independent respiratory wave time-stamped sequence to determine the independent snoring time-stamped sequence corresponding to each monitored object; and performing sleep monitoring analysis on the corresponding monitored objects based on the independent respiratory wave time-stamped sequence and / or the independent snoring time-stamped sequence of the same monitored object to obtain the sleep monitoring results corresponding to each monitored object. This achieves accurate sleep monitoring analysis of multiple monitored objects in the same space, resulting in precise sleep monitoring analysis results.
Owner:BEIJING XSMART CENTURY TECHNOLOGY CO LTD

Air conduction hearing aid far and near sound separation method

PendingCN121099244ADeaf-aid setsAir-conduction hearing aidBroadcasting
The invention relates to the field of hearing aids, in particular to an air conduction hearing aid far and near sound separation method, which comprises a hearing aid body module, a sound collection module, a control module, a broadcasting module, a vibration module and an isolation module, and is characterized in that the hearing aid body module comprises a body module and an ear hook module; the sound collection module comprises a stretching module and a sound receiver; the control module comprises a key module, a processing chip, a separation module and a transmission module; the hearing aid is worn, air is injected into the air bag, the air bag is expanded to isolate the upper broadcasting module and the lower broadcasting module and push the upper broadcasting module and the lower broadcasting module to abut against the ear canal, then the sound receiver extends out of the hearing aid body to enhance the receiving effect, and far and near sounds are separated and transmitted into the upper broadcasting module and the lower broadcasting module respectively. And the upper broadcasting module and the lower broadcasting module can synchronously start the upper vibration module and the lower vibration module during playing, so that a user can clearly hear the playing of far and near sounds and identify the far and near sounds.
Owner:LEFT POINT HEALTH IND (SHENZHEN) CO LTD

Method for adversarial training for universal sound separation

According to an aspect of the present disclosure there is provided a method for adversarial training of a separator (30) for universal sound separation of an audio mixture m of arbitrary sound sources Sk=1, . . . ,K, the method comprising: training a context-based discriminator (34) configured to provide a context-based loss cue based on a consideration of an input set of separated sound sources; and training the separator (30) to minimize a loss based on the context-based loss cue provided by the context-based discriminator (34); wherein training the context-based discriminator (34) comprises maximizing a loss based on a set of ground-truth sound sources and a fake set of separated sound sources, wherein the fake set of separated sound sources is sorted to match an order of the set of ground-truth sound sources, and wherein the fake set of separated sound sources comprises sources corresponding to separated sound sources estimated by the separator (30) and further comprises one or more ground-truth sound sources of the set of ground-truth sound sources.
Owner:DOLBY INTERNATIONAL AB

Dry and wet sound separation method and related device

The invention discloses a dry and wet sound separation method and a related device, and relates to the technical field of audio processing, the dry and wet sound state is introduced, and the signal state is determined, so that the separation strategy better fits the dynamic change of the signal. On the basis, when dry and wet sound separation is realized, a spatio-temporal information decoupling technology is adopted, spectrum information which changes rapidly and relatively stable spatial information are respectively processed and cooperatively utilized, interference of spatial aliasing on spectrum separation is avoided, a separation result is verified and optimized through spatial correlation, the problem that stereo spatial information is not fully utilized is solved, and the separation efficiency is improved. Therefore, the spatial information of stereophonic sound is fully mined, and the overall separation effect is improved.
Owner:IFLYTEK CO LTD

Vehicle external sound processing method and system and storage device

The invention provides a vehicle external sound processing method and system and a storage device. The method comprises the steps that original environment sound signals in multiple directions outside a vehicle are collected; based on a target sound pointing instruction, performing beam forming processing on the original environment sound signal, and extracting a first sound signal in a specific direction; performing noise suppression and target sound separation processing on the first sound signal to obtain a second sound signal; and performing high-fidelity audio reconstruction and amplification on the second sound signal to obtain an audio signal of the target sound, and outputting the audio signal to an in-vehicle loudspeaker. According to the method provided by the invention, directional screening and noise reduction of the environment sound outside the vehicle and high-fidelity playback in the cabin can be realized.
Owner:JINGDIAN AUTOMOTIVE ELECTRONICS (HUIZHOU) CO LTD

Convolution-augmented transformer models

Systems and methods can utilize a conformer model to process a data set for various data processing tasks, including, but not limited to, speech recognition, sound separation, protein synthesis determination, video or other image set analysis, and natural language processing. The conformer model can use feed-forward blocks, a self-attention block, and a convolution block to process data to learn global interactions and relative-offset-based local correlations of the input data.
Owner:GOOGLE LLC

Sound separation method, apparatus, electronic device, and storage medium

The present application relates to the technical field of sound processing, and provides a sound separation method, device, electronic equipment and storage medium, the method comprising: obtaining a first sound signal and a second sound signal; determining the short-time energy value of the ambient sound and the coherent sound respectively, and the difference factor of the coherent sound on different channels, based on the short-time energy value of the first sound signal and the second sound signal respectively, and the linear combination relationship between the ambient sound and the coherent sound contained in the first sound signal and the second sound signal respectively; mapping the short-time energy value of the ambient sound and the coherent sound respectively, and the difference factor, to the sound separation weight based on the weight mapping relationship, the weight mapping relationship being obtained by linear fitting based on the linear combination relationship; separating the ambient sound and the coherent sound from the first sound signal and the second sound signal based on the sound separation weight. The method, device, electronic equipment and storage medium provided by the present application ensure the effectiveness and reliability of the separation of the ambient sound and the coherent sound.
Owner:IFLYTEK (SUZHOU) TECH CO LTD

Voice consultation device, voice consultation method, and storage medium

The application relates to a voice consultation device, a voice consultation method and a storage medium, and belongs to the technical field of computers.The device comprises a microphone module, an AEC circuit, a camera and a processor; the AEC circuit separates the sound of the loudspeaker of the voice consultation device, and does not send the sound to voice recognition; the voice consultation device can improve the clarity of the audio device collected from the source; meanwhile, the angle of the user facing the microphone is analyzed through a beamforming algorithm, so that the voice signals outside the angle are suppressed, the multichannel human voice signals inside the angle are integrated into single-channel audio first, and the single-channel audio is sent to voice recognition, the clarity of the audio data can be further improved, and the accuracy of voice recognition is improved. In addition, the distance between the user and the voice consultation device is determined through the camera, and the recognition mode corresponding to the distance is automatically switched, so that the accuracy of voice recognition is improved, and the response accuracy of the voice consultation device is improved.
Owner:AISPEECH CO LTD

Stereo processing method and device

The invention provides a stereophonic sound processing method and device, and relates to the technical field of sound processing, and the method comprises the steps: carrying out the dry and wet sound separation processing of a stereophonic sound audio signal comprising a left sound channel signal and a right sound channel signal, and obtaining a dry sound, a left channel wet sound, and a right channel wet sound; wherein the left channel wet sound is a signal component irrelevant to a right channel signal in the left channel signal, the right channel wet sound is a signal component irrelevant to the left channel signal in the right channel signal, and the dry sound is a signal component with the same amplitude and phase in the left channel signal and the right channel signal; performing filtering processing on the dry sound, the left channel wet sound and the right channel wet sound by using a preset first control filter, a preset second control filter and a preset third control filter respectively to obtain three paths of corresponding filtering signals; and superposing the three paths of filtering signals to generate a multi-channel output signal for driving the loudspeaker array.
Owner:IFLYTEK (SUZHOU) TECH CO LTD

Privacy-respecting detection and localization of sounds in autonomous driving applications

The described aspects and implementations enable privacy-respecting detection, separation, and localization of sounds in vehicle environments. The techniques include obtaining, using audio detector(s) of a vehicle, a sound recording that includes a plurality of elemental sounds (ESs) in a driving environment of the vehicle, and processing, using a sound separation model, the sound recording to separate individual ESs of the plurality of ESs. The techniques further include identifying a content of individual ESs and causing a driving path of the vehicle to be modified in view of the identified content of the individual ESs. Further techniques include rendering speech imperceptibly by redacting temporal portions of the speech, using sound recognition models to identify and discard recordings of speech, and driving at speeds that exceed threshold speeds at which speech becomes imperceptible from noise masking.
Owner:WAYMO LLC

A single-channel sound separation method

ActiveCN115331693BSpeech analysisNonnegative matrix factorisationNonnegative matrix
The present invention provides a single-channel sound separation method, belonging to the field of single-channel mixed sound signal separation. The method comprises: mixing a first single-channel target sound signal and a first single-channel interference sound signal to obtain a first single-channel mixed sound signal; constructing a sound separation network based on a semi-nonnegative matrix factorization algorithm, training the sound separation network with the first single-channel target sound signal and the first single-channel mixed sound signal to obtain a sound separation model; and inputting a second single-channel mixed sound signal into the sound separation model for separation to obtain a second target sound signal. The present invention enables the network to have a better ability to extract target sound signal components, thereby better reconstructing the target sound and achieving a better sound separation effect.
Owner:GUANGZHOU INST OF RAILWAY TECH