Audio Personalisation Using Smoothed HRTFs for 3D Immersion
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing methods for achieving realistic audio immersion in media content, such as videogames, are complex and require specialist equipment, making it impractical to replicate the unique audio interactions of individual users' heads for personalized 3D audio experiences.
Innovation Solution
An audio personalization method and system that captures head-related transfer functions (HRTFs) for users by matching their head morphology with reference individuals or through calibration tests, and uses smoothed HRTFs to simulate large or volumetric sound sources, allowing for the reproduction of 3D audio directional sources and environments.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If techniques for achieving realistic audio immersion are used, then audio realism is improved, but device complexity increases and specialist equipment is required
Solution Approach 1:
The patent creates a simplified copy of the complex HRTF measurement process by using image analysis of head morphology to generate approximate HRTFs. Instead of requiring actual acoustic measurements with specialized equipment, the system copies the essential characteristics needed for spatial audio by analyzing visual data of the user's head and ears, thereby achieving realistic audio immersion without complex measurement devices
Solution Approach 2:
The patent replaces the mechanical/acoustic measurement system with an optical/image processing system. Instead of using microphones, acoustic chambers, and physical HRTF measurement equipment, the system uses image capture devices and computer vision algorithms to analyze head morphology and generate HRTFs, substituting a simpler optical system for the complex acoustic measurement system
2Adaptability or versatility
If HRTF measurements are performed for each user, then audio personalisation is improved, but time and resources are consumed
Solution Approach 1:
The patent performs preliminary action by pre-computing and storing HRTFs for multiple reference individuals with different head morphologies. When a new user is encountered, the system quickly matches the user's head morphology to one of the pre-computed reference HRTFs, avoiding the need for time-consuming real-time measurements. This preliminary preparation of reference data enables rapid personalization
Solution Approach 2:
The patent creates a universal system where a single set of reference HRTFs can serve multiple users with similar head morphologies. By categorizing and storing HRTFs for reference individuals representing different head types, the system can universally apply the appropriate reference HRTF to any user by matching their morphology, making the personalization process efficient and scalable across many users
Data Source
Figure 1
Figure 2A~2B
Figure 3A~3B
AI summary
An audio personalisation method for a user, to reproduce an area-based or volumetric sound source, comprised the steps of, for a head related transfer function 'HRTF' associated with the user, smoothing HRTF coefficients relating to peaks and notches in the HRTF's spectral response, responsive to the size of the area or volume of the sound source; filtering the sound source using the smoothed HRTF for the notional position of the sound source; and outputting the filtered sound source signal for playback to the user.