Dynamic Multi-Channel Audio Generation from Subject Perspective

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional multi-channel audio systems require excessive human intervention and are labor-intensive, failing to provide an immersive audio experience by reproducing a specific user's perspective, such as a sports player's, due to pre-defined settings and lack of dynamic audio generation.

Innovation Solution

A media content packaging and distribution system that dynamically generates multi-channel audio by selecting a subject-of-interest and corresponding audio-capture devices based on location, social media trends, and user preferences, using a server to mix and encode audio streams into a surround sound environment simulating the acoustic experience from the subject's perspective.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Manufacturing precision

If human operators manually assign audio streams to channels, then audio channel assignment can be precisely controlled, but the process becomes time-intensive and labor-intensive

Engineering Contradiction:
Improveaudio channel assignment precisionVSAvoidaudio generation efficiency
Core Design Contradiction:
Manufacturing precisionVSProductivity

Solution Approach 1:

The system enables automatic audio stream assignment by allowing the audio streams themselves to carry metadata identifiers that automatically map to appropriate audio channels, eliminating the need for manual operator intervention while maintaining precise channel assignment

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

Audio streams are pre-tagged with metadata identifiers during recording or encoding, so that when the multi-channel audio is generated, the assignment to specific channels has already been determined in advance, removing the need for real-time manual assignment

Inventive Principle:
Principle #10Preliminary action

2Ease of operation

If pre-defined settings are used for audio routing, then the system configuration is simple, but the audio experience lacks immersion and cannot reproduce specific user perspectives

Engineering Contradiction:
Improvesystem configuration simplicityVSAvoidaudio perspective customization
Core Design Contradiction:
Ease of operationVSAdaptability or versatility

Solution Approach 1:

The system transitions from static pre-defined audio routing to dynamic audio generation where the audio perspective can change based on the selected subject-of-interest, allowing the audio environment to adapt in real-time while maintaining simple system operation through automated processing

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The system changes audio parameters such as spatial positioning, volume levels, and channel distribution based on the identified subject-of-interest, enabling different audio perspectives without requiring complex manual reconfiguration

Inventive Principle:
Principle #35Parameter changes

3Reliability

If multiple audio-capture devices are used to capture audio from different subjects, then audio coverage is comprehensive, but selecting and mixing the appropriate audio streams becomes complex

Engineering Contradiction:
Improveaudio capture completenessVSAvoidaudio stream selection complexity
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

The system extracts and isolates the audio streams associated with the selected subject-of-interest from the multiple captured audio streams, separating the relevant audio data from the rest while maintaining comprehensive audio coverage capability

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

Metadata identifiers act as intermediaries between the multiple audio-capture devices and the audio mixing process, providing a simple mechanism to identify and select the appropriate audio streams without complex selection logic

Inventive Principle:
Principle #24Intermediary (Mediator)

Data Source

PatentUS10341762B2Dynamic generation and distribution of multi-channel audio from the perspective of a specific subject of interest
Publication Date: 2019.07.02 SONY GROUP CORP
  • US10341762B2 patent drawing
  • US10341762B2 patent drawing
  • US10341762B2 patent drawing

AI summary

A media content packaging and distribution system for dynamic generation of multi-channel audio includes a server, which stores location information of a plurality of subjects located in a defined area. A subject-of-interest is selected from the plurality of subjects in the defined area. Thereafter, a set of audio-capture devices are selected from the plurality of audio-capture devices. A set of audio streams are received from the selected set of audio-capture devices. A multi-channel audio is generated based on the received set of audio streams. The generated multi-channel audio is communicated to a consumer device. Based on an output of the multi-channel audio by the consumer device, an acoustic environment is reproduced as a surround sound environment at the consumer device from a perspective of the subject-of interest.