Enhanced Sound Field Description for 6DoF Reproduction
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current Ambisonics and DirAC sound field representations lack sufficient information to enable six-degrees-of-freedom (6DoF) reproduction, as they do not provide translational shift capabilities, which are necessary for interactive and immersive audio experiences in virtual reality applications.
Innovation Solution
An enhanced sound field description is generated by incorporating meta data that includes spatial information, such as distance information, allowing for the calculation of a modified sound field at a different reference location, enabling 6DoF reproduction by translating the listener's position within the sound scene.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If traditional Ambisonics or DirAC sound field representations are used, then the sound field can be represented with a limited number of signals, but the system lacks translational shift capabilities necessary for 6DoF reproduction
Solution Approach 1:
The patent applies preliminary action by pre-calculating and storing transfer functions for translating sound field descriptions between different reference locations. These transfer functions are computed in advance based on the known positions of sound sources and reference locations, allowing rapid 6DoF reproduction without real-time complex calculations. This enables the system to provide translational shift capabilities while maintaining efficient signal representation.
2Measurement precision
If more signals are used to enhance spatial resolution and listener sweet-spot area, then the sound field description becomes more accurate, but the data transmission and processing complexity increases
Solution Approach 1:
The patent changes parameters by introducing transfer functions that depend on the relative positions between reference locations and sound sources. Instead of increasing the number of Ambisonics signals to achieve translational capability, the system modifies the existing sound field description by applying position-dependent transfer functions. This approach maintains spatial resolution while avoiding the complexity increase associated with higher-order Ambisonics signals.
Data Source
AI summary
An apparatus for generating an enhanced sound field description has: a sound field generator for generating at least two sound field layer descriptions indicating sound fields with respect to at least one reference location; and a meta data generator for generating meta data relating to spatial information of the sound fields, wherein the sound field descriptions and the meta data constitute the enhanced sound field description. The meta data can be a geometrical information for each layer such as a representative distance to the reference location.


