Network Media Enhancement Server Personalized Audio Processing

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current telecommunications networks lack personalized audio processing capabilities to effectively enhance speech intelligibility for individuals with hearing impairments and those in noisy environments, as existing solutions are expensive, limited, and not adaptable to individual hearing needs.

Innovation Solution

A network-based media enhancement system that uses a media enhancement server to apply personalized audio processing parameters, such as automatic gain control and noise cancellation, based on unique identifiers associated with specific listeners or devices, to improve speech comprehension and comfort during calls.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If network-based audio processing is implemented, then speech intelligibility is improved, but device complexity increases

Engineering Contradiction:
Improvespeech intelligibilityVSAvoiddevice complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The patent introduces a network-based media enhancement function as an intermediary component between the telephony network and the end user device. This separate functional element processes audio signals through techniques like acoustic echo cancellation, noise cancellation, and speech enhancement, thereby improving speech intelligibility without adding complexity to the end user's device. The complex processing is centralized in the network infrastructure rather than distributed to individual devices.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Measurement precision

If personalized audio processing parameters are applied, then hearing acuity is improved, but adaptability requirements increase

Engineering Contradiction:
Improvehearing acuityVSAvoidadaptability requirements
Core Design Contradiction:
Measurement precisionVSAdaptability or versatility

Solution Approach 1:

The patent implements preliminary action by pre-determining personalized audio processing parameters for each user through hearing evaluations and assessments. These parameters, including gain control settings, frequency compensation, and noise cancellation profiles, are established in advance and stored in a database. When a user initiates a call, the pre-configured parameters are automatically retrieved and applied, eliminating the need for real-time adaptation during conversations and simplifying the user experience.

Inventive Principle:
Principle #10Preliminary action

3Reliability

If existing specialty telephones are used, then hearing impairment support is provided, but cost increases

Engineering Contradiction:
Improvehearing impairment supportVSAvoidcost
Core Design Contradiction:
ReliabilityVSQuantity of substance

Solution Approach 1:

The patent transforms the audio enhancement capability from a specialized hardware feature available only in expensive specialty telephones into a universal network service. The media enhancement function can serve all users regardless of their specific hearing needs, providing acoustic echo cancellation, noise cancellation, and speech enhancement to everyone. Personalized parameters are applied selectively based on user profiles, allowing the system to support hearing-impaired individuals effectively while making the technology accessible to the general population through standard telephony infrastructure.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Data Source

PatentUS9020621B1Network based media enhancement function based on an identifier
Publication Date: 2015.04.28 COCHLEAR LIMITED
  • US9020621B1 patent drawing
  • US9020621B1 patent drawing
  • US9020621B1 patent drawing

AI summary

A network based processing element for processing audio information improves the understanding of speech or music for intended listeners based on an identifier. The processing involves performing a media enhancement function, where a parameter affecting the utilization or performance aspects of the media enhancement function are dependent upon the identifier. A “media enhancement server” (MES) is included, whereby the audio of a telephone call, video call, multimedia program or other stream to be heard by a specific listener is processed using a personalized audio enhancement parameter to enhance the audio signal such that the listener will enjoy a benefit, such as better comprehension of the information, reduced listening effort, and more listening comfort during the call. The personalized parameters are stored and retrieved based upon the identifier, and used within the MES. The audio portion of the call or stream could be speech, music, or a combination.