Spatial Audio Rendering via Object Metadata

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current audio content delivery systems provide only pre-mixed stereo audio, failing to offer a user-customized immersive experience that simulates the spatial features of a venue, limiting the ability to recreate the exact audio environment.

Innovation Solution

A computer system generates audio files and metadata with spatial features for objects at a venue, allowing electronic devices to render audio files based on these features, creating a user-customized immersive audio experience by simulating the audio environment.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If pre-mixed stereo audio content is provided, then the audio content delivery is simple and efficient, but the user cannot experience the spatial features of the venue

Engineering Contradiction:
Improvespatial experience customizationVSAvoidaudio processing complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent segments the audio content into separate audio files for each object at the venue, rather than providing pre-mixed stereo audio. Each audio file is associated with metadata containing spatial features, allowing the electronic device to independently process and render each object's audio with its specific spatial characteristics, thereby enabling spatial experience customization without requiring complex pre-processing of the entire audio mix

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent applies preliminary action by pre-generating audio files and their associated metadata with spatial features at the venue before transmission to the electronic device. The spatial features (position, direction, distance) are predetermined and embedded in the metadata, allowing the device to directly render the spatial audio experience without requiring complex real-time processing or user configuration

Inventive Principle:
Principle #10Preliminary action

2Adaptability or versatility

If completed audio content is transmitted, then the transmission efficiency is high, but the user cannot customize the audio experience

Engineering Contradiction:
Improveaudio experience customizationVSAvoiddata transmission volume
Core Design Contradiction:
Adaptability or versatilityVSQuantity of substance

Solution Approach 1:

The patent extracts the spatial features from the completed audio content and separates them into independent metadata associated with each audio file. Instead of transmitting a single pre-mixed stereo audio stream, the system transmits multiple audio files with their corresponding spatial metadata, allowing the electronic device to reconstruct the audio experience with customizable spatial characteristics while managing data transmission through efficient metadata compression

Inventive Principle:
Principle #2Taking out (Extraction)

Data Source

PatentUS11930348B2Computer system for realizing customized being-there in association with audio and method thereof
Publication Date: 2024.03.12 NAVER CORP
  • US11930348B2 patent drawing
  • US11930348B2 patent drawing
  • US11930348B2 patent drawing

AI summary

A method by a computer system including generating audio files based on respective audio signals, the audio signals having been respectively generated from a plurality of objects at a venue, generating metadata including spatial features at the venue that are respectively set for the objects, and transmitting the audio files and the metadata for the objects to a first electronic device to cause the first electronic device to realize a being-there at the venue by rendering the audio files based on the spatial features in the metadata may be provided.