The invention relates to the technical field of self-adaptive spatial audio rendering methods, in particular to a self-adaptive spatial audio rendering method based on
head tracking, which comprises the following steps of: S1, acquiring the head orientation of a user based on
head tracking equipment, and calculating a current binaural
transfer function of the user; s2, based on a deep neural network, training a binaural
transfer function for each user according to the user feature information, and generating a personalized binaural
transfer function and an audio rendering parameter; s3, inputting spatial audio data, and rendering the spatial audio data into left and right sound channel audio signals; s4, performing audio time
delay correction on the rendered left and right sound channel audio signals, and outputting the rendered sound effect; s5, the audio signals are input into the TWS earphone through the
ear canal model, spatial audio rendering is completed, and by means of the
deep learning model and the optimized
signal processing algorithm, high-quality audio output is successfully guaranteed, and meanwhile the
millisecond-level response speed is achieved.