The invention discloses a training method of a multi-sound-source direction-of-arrival
estimation model, a multi-sound-source direction-of-arrival
estimation method, equipment, a medium and a product, and relates to the technical field of
signal processing and
artificial intelligence crossing, and the training method comprises the steps: calculating the time-frequency characteristics and cross-correlation characteristics of original signals of a multi-channel array; respectively carrying out position coding on the time-frequency characteristics and the
microphone position information; fusing the features into an input matrix, and inputting the input matrix into a
backbone network of an improved Transform model; calculating the attention
score of the input matrix and generating head output by using the first multi-head self-attention layer, inputting the head output into the second multi-head self-attention layer, calculating the attention
score and generating head output, and inputting the output into a multi-task output module to obtain a direction
estimation result. The multi-head attention mechanism is introduced, long-time and multi-band complex dependence is captured, the sound source distinguishing capacity is improved, multiple heads capture
direction information from different view angles, and it is guaranteed that high accuracy can still be guaranteed in the complex environment.