The application discloses a subcutaneous
speech enhancement method based on SATF
estimation and
stream matching, which constructs a
simulation adaptive framework containing SATF
estimation and
stream matching
noise simulation double branches, performs normalized cross-correlation pre-alignment on a subcutaneous speech pair and executes an STFT transformation, constructs an SATF
library through STFT domain least square
estimation and frequency
smoothing, and synthesizes band-pass channel
colored simulation speech based on the SATF
library and
dynamic noise; scene-related structured
residual noise is generated by combining
stream matching, simulation speech and structured
residual noise are fused to obtain final simulation noisy speech, and the final simulation noisy speech is paired with clean speech to generate aligned
simulation training speech pairs, and a pre-training
speech enhancement model is adaptively trained by using the training speech pairs. The application can effectively alleviate the problems of subcutaneous speech data scarcity and alignment deviation, improve
speech enhancement effect, and adapt to applications of full-implantable cochlear implants and other implantable hearing systems.