Multichannel Downmix Audio Object Coding Separation

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing audio object coding methods fail to effectively decode multiple objects from an encoded multi-object signal using multichannel downmix and additional control data, particularly when dealing with more than one downmix channel, limiting the separation quality and flexibility of object rendering.

Innovation Solution

A method that jointly decodes multiple downmix channels using a spatial audio object coder, which generates downmix information and object parameters to reconstruct audio objects, allowing for flexible scaling of separation quality and incorporating correlation between objects, enabling the use of multiple downmix channels for improved object separation.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Device complexity

If single-channel downmix coding is used, then device complexity is reduced, but separation quality of audio objects deteriorates

Engineering Contradiction:
Improvecoding complexityVSAvoidseparation quality
Core Design Contradiction:
Device complexityVSMeasurement precision

Solution Approach 1:

The patent transitions from single-channel downmix coding to multichannel downmix coding, adding spatial dimensionality to the coding process. By distributing audio objects across multiple downmix channels and using multichannel coherence parameters, the system achieves better separation quality while maintaining manageable complexity through efficient parameter representation.

Inventive Principle:
Principle #17Another dimension (Dimensionality change)

2Measurement precision

If multichannel downmix coding is used, then separation quality of audio objects is improved, but device complexity increases

Engineering Contradiction:
Improveseparation qualityVSAvoidcoding complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The patent introduces multichannel coherence parameters and object-specific downmix parameters as new parameter types that capture spatial relationships between multiple downmix channels. These parameter changes enable the decoder to separate audio objects effectively from multichannel downmixes without requiring complex processing, thus improving separation quality while controlling device complexity.

Inventive Principle:
Principle #35Parameter changes

3Adaptability or versatility

If existing audio object coding methods are used, then compatibility with single-channel devices is maintained, but flexibility in rendering configurations is limited

Engineering Contradiction:
Improverendering flexibilityVSAvoidcompatibility
Core Design Contradiction:
Adaptability or versatilityVSReliability

Solution Approach 1:

The patent creates a universal coding framework that can represent audio objects in multichannel downmixes while maintaining compatibility with single-channel playback. The encoder can adaptively choose to encode objects into single-channel or multichannel downmixes based on the capabilities of the playback device, making the system universally applicable across different device types while providing enhanced flexibility for multichannel rendering when available.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Data Source

PatentUS9565509B2Enhanced coding and parameter representation of multichannel downmixed object coding
Publication Date: 2017.02.07 DOLBY INTERNATIONAL AB
  • US9565509B2 patent drawing
  • US9565509B2 patent drawing
  • US9565509B2 patent drawing

AI summary

An audio object coder for generating an encoded object signal using a plurality of audio objects includes a downmix information generator for generating downmix information indicating a distribution of the plurality of audio objects into at least two downmix channels, an audio object parameter generator for generating object parameters for the audio objects, and an output interface for generating the imported audio output signal using the downmix information and the object parameters. An audio synthesizer uses the downmix information for generating output data usable for creating a plurality of output channels of the predefined audio output configuration.