This disclosure provides methods, devices, and systems for
audio signal mixing. The present implementations more specifically relate to mixing audio signals from a
microphone array by performing fixed
beamforming to generate beams, reducing
noise on the beams, and mixing the beams to generate a final
audio signal for playback. In some aspects, an
audio mixing system includes a fixed beamformer to generate beams from audio signals from a
microphone array and
noise reduction units (NRUs) to reduce a
noise component of each audio beam. The
system also includes logic to calculate a
signal characteristic of each reduced noise audio beam to determine, based on the
signal characteristics, the reduced noise audio beams that include a speech component. The logic also generates a
gain for each audio beam based on the selection, with the gains used in beam mixing. In some aspects, the NRU includes a neural network
noise reduction unit.