Spatial Audio Rendering via Object Metadata
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current audio content delivery systems provide only pre-mixed stereo audio, failing to offer a user-customized immersive experience that simulates the spatial features of a venue, limiting the ability to recreate the exact audio environment.
Innovation Solution
A computer system generates audio files and metadata with spatial features for objects at a venue, allowing electronic devices to render audio files based on these features, creating a user-customized immersive audio experience by simulating the audio environment.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If pre-mixed stereo audio content is provided, then the audio content delivery is simple and efficient, but the user cannot experience the spatial features of the venue
Solution Approach 1:
The patent segments the audio content into separate audio files for each object at the venue, rather than providing pre-mixed stereo audio. Each audio file is associated with metadata containing spatial features, allowing the electronic device to independently process and render each object's audio with its specific spatial characteristics, thereby enabling spatial experience customization without requiring complex pre-processing of the entire audio mix
Solution Approach 2:
The patent applies preliminary action by pre-generating audio files and their associated metadata with spatial features at the venue before transmission to the electronic device. The spatial features (position, direction, distance) are predetermined and embedded in the metadata, allowing the device to directly render the spatial audio experience without requiring complex real-time processing or user configuration
2Adaptability or versatility
If completed audio content is transmitted, then the transmission efficiency is high, but the user cannot customize the audio experience
Solution Approach 1:
The patent extracts the spatial features from the completed audio content and separates them into independent metadata associated with each audio file. Instead of transmitting a single pre-mixed stereo audio stream, the system transmits multiple audio files with their corresponding spatial metadata, allowing the electronic device to reconstruct the audio experience with customizable spatial characteristics while managing data transmission through efficient metadata compression
Data Source
AI summary
A method by a computer system including generating audio files based on respective audio signals, the audio signals having been respectively generated from a plurality of objects at a venue, generating metadata including spatial features at the venue that are respectively set for the objects, and transmitting the audio files and the metadata for the objects to a first electronic device to cause the first electronic device to realize a being-there at the venue by rendering the audio files based on the spatial features in the metadata may be provided.


