Dialog Loudness Normalization for Consistent Program Audio

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current audio processing systems lack a standardized method for maintaining consistent loudness levels across different programs and advertisements, leading to unexpected volume changes when switching between channels or programs, due to mismatched dynamic ranges and improper setting of the Dialog Normalization (DIALNORM) parameter.

Innovation Solution

A method and system for correcting audio levels by retrieving stored program assets, identifying dialog, determining the loudness using psycho-acoustic criteria, and re-encoding the audio if the initial loudness setting differs significantly from the determined loudness, ensuring consistent loudness settings across programs and advertisements.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If different programs and advertisements are encoded with their original loudness settings, then each program maintains its intended audio characteristics, but viewers experience unexpected volume changes when switching between programs

Engineering Contradiction:
Improveaudio characteristicsVSAvoidvolume adjustment
Core Design Contradiction:
Adaptability or versatilityVSEase of operation

Solution Approach 1:

The patent applies parameter changes by modifying the loudness encoding parameter (DIALNORM) of audio content. The system analyzes the original loudness characteristics of programs and advertisements, then re-encodes them with standardized loudness parameters to ensure consistent playback volume across different content sources, eliminating the need for manual viewer adjustments.

Inventive Principle:
Principle #35Parameter changes

2Ease of operation

If audio is re-encoded to standardize loudness levels, then consistent volume is achieved across programs, but processing time and computational resources increase

Engineering Contradiction:
Improvevolume consistencyVSAvoidprocessing time
Core Design Contradiction:
Ease of operationVSLoss of time

Solution Approach 1:

The patent implements preliminary action by pre-analyzing and pre-standardizing the loudness parameters of audio content during off-peak hours or in the background. The system measures the original loudness characteristics, determines appropriate normalization parameters, and prepares standardized audio versions in advance, so that when viewers request content, the loudness normalization is already complete and no real-time processing is required.

Inventive Principle:
Principle #10Preliminary action

3Ease of operation

If DIALNORM parameter is properly set during encoding, then loudness consistency is maintained, but additional encoding steps are required

Engineering Contradiction:
Improveloudness consistencyVSAvoidencoding process
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The patent applies self-service by implementing an automated system that measures the original loudness characteristics of audio content, automatically determines the appropriate DIALNORM parameter values, and performs re-encoding without manual intervention. The system includes built-in loudness measurement capabilities and automatic parameter adjustment algorithms that adapt to different content types (programs, advertisements, sports events), eliminating the need for manual encoding configuration.

Inventive Principle:
Principle #25Self-service

Data Source

PatentUS8379880B2Methods and systems for determining audio loudness levels in programming
Publication Date: 2013.02.19 TIME WARNER CABLE ENTERPRISES LLC
  • US8379880B2 patent drawing
  • US8379880B2 patent drawing
  • US8379880B2 patent drawing

AI summary

An example of a method of correcting an audio level of a stored program asset comprises retrieving a stored program asset having audio encoded at a first loudness setting. Dialog of the audio of the asset is identified, a loudness of the dialog is determined and the determined loudness is compared to the first loudness setting. The asset is re-encoded at a second loudness setting corresponding to the determined loudness, if the first loudness setting and the second loudness are different by more than a predetermined amount. The determined loudness is preferably a DIALNORM of the dialog. The asset may be stored with the re-encoded loudness setting. The method may be applied to programs as they are being received from a source, as well. Aspects of the method may also be applied to programs to be provided by a source. Systems are also disclosed.