Audio Data Conversion With Metadata for Bit-Depth Reduction

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing audio signal processing technologies do not adequately enhance the quality of sound data, particularly in converting between different bit formats and incorporating accessory information related to the sound collection device or sound source.

Innovation Solution

A method and apparatus that acquires first sound data in a floating point format, attaches accessory information including device or sound source information, and edits the data to create second sound data with a smaller bit number, while incorporating video data imaging and editing steps to improve sound quality.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Quantity of substance

If sound data is converted from floating point format to fixed point format with smaller bit number, then data size is reduced and ease of storage is improved, but sound quality and dynamic range deteriorate

Engineering Contradiction:
Improvedata sizeVSAvoidsound quality
Core Design Contradiction:
Quantity of substanceVSManufacturing precision

Solution Approach 1:

The patent applies preliminary action by performing dither noise addition and noise shaping processing before the quantization conversion from floating point to fixed point format. This preprocessing prepares the sound data in advance to minimize quality loss during the subsequent bit reduction, allowing the system to achieve both compact storage and preserved sound quality.

Inventive Principle:
Principle #10Preliminary action

2Loss of information

If accessory information is added to sound data, then information completeness is improved, but data complexity and processing difficulty increase

Engineering Contradiction:
Improveinformation completenessVSAvoiddata complexity
Core Design Contradiction:
Loss of informationVSDevice complexity

Solution Approach 1:

The patent merges accessory information (such as device information, sound source information, and processing information) with the sound data into a single integrated data structure. This combining approach ensures that all relevant information is preserved and readily accessible without requiring separate management systems, thus improving information completeness while avoiding excessive complexity.

Inventive Principle:
Principle #5Merging (Combining)

3Speed

If down-sampling is performed to reduce frequency, then processing speed is improved, but sound quality deteriorates

Engineering Contradiction:
Improveprocessing speedVSAvoidsound quality
Core Design Contradiction:
SpeedVSManufacturing precision

Solution Approach 1:

The patent applies preliminary action by performing anti-aliasing filtering and dither noise addition before down-sampling the sound data. This preprocessing ensures that the frequency reduction is performed in a way that minimizes quality loss, allowing the system to achieve faster processing speeds while preserving sound quality through proper preparation of the data beforehand.

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS20250266050A1Creation method and creation apparatus
Publication Date: 2025.08.21 FUJIFILM CORP
  • US20250266050A1 patent drawing
  • US20250266050A1 patent drawing
  • US20250266050A1 patent drawing

AI summary

A creation method of the present disclosure includes an acquisition step of acquiring first sound data of a floating point format, based on a sound that is generated from a sound source and is collected by a sound collection device, and an accessory information creation step of creating accessory information that is attached to the first sound data and includes device information, which relates to the sound collection device, or sound source information, which relates to the sound source.