Audio Data Conversion With Metadata for Bit-Depth Reduction
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing audio signal processing technologies do not adequately enhance the quality of sound data, particularly in converting between different bit formats and incorporating accessory information related to the sound collection device or sound source.
Innovation Solution
A method and apparatus that acquires first sound data in a floating point format, attaches accessory information including device or sound source information, and edits the data to create second sound data with a smaller bit number, while incorporating video data imaging and editing steps to improve sound quality.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Quantity of substance
If sound data is converted from floating point format to fixed point format with smaller bit number, then data size is reduced and ease of storage is improved, but sound quality and dynamic range deteriorate
Solution Approach 1:
The patent applies preliminary action by performing dither noise addition and noise shaping processing before the quantization conversion from floating point to fixed point format. This preprocessing prepares the sound data in advance to minimize quality loss during the subsequent bit reduction, allowing the system to achieve both compact storage and preserved sound quality.
2Loss of information
If accessory information is added to sound data, then information completeness is improved, but data complexity and processing difficulty increase
Solution Approach 1:
The patent merges accessory information (such as device information, sound source information, and processing information) with the sound data into a single integrated data structure. This combining approach ensures that all relevant information is preserved and readily accessible without requiring separate management systems, thus improving information completeness while avoiding excessive complexity.
3Speed
If down-sampling is performed to reduce frequency, then processing speed is improved, but sound quality deteriorates
Solution Approach 1:
The patent applies preliminary action by performing anti-aliasing filtering and dither noise addition before down-sampling the sound data. This preprocessing ensures that the frequency reduction is performed in a way that minimizes quality loss, allowing the system to achieve faster processing speeds while preserving sound quality through proper preparation of the data beforehand.
Data Source
AI summary
A creation method of the present disclosure includes an acquisition step of acquiring first sound data of a floating point format, based on a sound that is generated from a sound source and is collected by a sound collection device, and an accessory information creation step of creating accessory information that is attached to the first sound data and includes device information, which relates to the sound collection device, or sound source information, which relates to the sound source.


