MPEG-H 3D Audio Position Compensation for 3DOF+ Head Movement

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

The MPEG-H 3D Audio standard does not support small translational movements of a user's head, limiting the immersive experience to rotational head movements only, which results in incorrect perception of audio objects close to the listener during translational head movements.

Innovation Solution

A method and apparatus that modify audio object positions based on listener displacement and orientation information, applying translations and rotational transformations to account for small translational and rotational head movements, ensuring audio objects are perceived from a fixed position relative to the listener's head, enhancing the immersive experience.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If the MPEG-H 3D Audio standard functionality for rotational head movements is used, then the audio scene can remain spatially stationary under change of listener's head orientation, but small translational movements of the user's head cannot be accounted for

Engineering Contradiction:
Improvesupport for head movementsVSAvoidaccuracy of audio object perception
Core Design Contradiction:
Adaptability or versatilityVSReliability

Solution Approach 1:

The patent extends the static 3 DoF audio scene to a dynamic 3 DoF+ environment by introducing translational movement compensation. The audio object positions are dynamically adjusted based on detected head displacement, allowing the system to adapt to both rotational and translational head movements while maintaining accurate spatial perception

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The patent enhances the MPEG-H 3D Audio standard by adding translational movement handling to the existing rotational movement functionality. The unified processing approach handles both 3 DoF (rotational) and 3 DoF+ (translational) scenarios through a single extended framework that works with existing audio objects and position information

Inventive Principle:
Principle #6Universality (Multi-functionality)

2Ease of operation

If the audio scene remains spatially stationary under change of listener's head orientation, then rotational head movements are supported, but translational head movements result in incorrect perception of audio objects

Engineering Contradiction:
Improverotational head movement supportVSAvoidaudio object position accuracy
Core Design Contradiction:
Ease of operationVSManufacturing precision

Solution Approach 1:

The patent modifies the audio object position parameters by applying translational offsets based on detected head displacement. The position information is adjusted using displacement vectors derived from sensor data, transforming the stationary audio scene into a dynamically positioned one that compensates for head translation while preserving rotational characteristics

Inventive Principle:
Principle #35Parameter changes

Data Source

PatentUS12395810B2Methods, apparatus and systems for three degrees of freedom (3DOF+) extension of MPEG-H 3D audio
Publication Date: 2025.08.19 DOLBY INTERNATIONAL AB
  • US12395810B2 patent drawing
  • US12395810B2 patent drawing
  • US12395810B2 patent drawing

AI summary

Described is a method of processing position information indicative of an object position of an audio object, wherein the object position is usable for rendering of the audio object, that comprises: obtaining listener orientation information indicative of an orientation of a listener's head; obtaining listener displacement information indicative of a displacement of the listener's head; determining the object position from the position information; modifying the object position based on the listener displacement information by applying a translation to the object position; and further modifying the modified object position based on the listener orientation information. Further described is a corresponding apparatus for processing position information indicative of an object position of an audio object, wherein the object position is usable for rendering of the audio object.