Audio Feedback Time Shift Filter for Voice Recognition

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Vocal command recognition systems face challenges in distinguishing user commands from audio feedback and environmental noise, particularly when program audio content, such as lyrics, interferes with voice recognition algorithms.

Innovation Solution

An audio feedback time shift filter system that separates program audio feedback from environmental audio using a deterministic delay to synchronize and filter out program content audio feedback, preventing misinterpretation of user commands by compensating for feedback loop delays.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If vocal command recognition is implemented to control audio systems, then user convenience and hands-free operation are improved, but audio feedback from program content causes misinterpretation and false commands

Engineering Contradiction:
Improveuser convenienceVSAvoidcommand recognition accuracy
Core Design Contradiction:
Ease of operationVSReliability

Solution Approach 1:

The audio signal is segmented into program content audio and environmental audio components. The system separates these segments by identifying and removing the program content audio portion from the combined audio signal, allowing pure environmental audio (user commands) to be processed for recognition.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The harmful program content audio feedback is extracted and removed from the environmental audio signal. The system identifies the program content audio patterns and extracts them from the mixed audio signal before voice recognition processing, preventing false command generation.

Inventive Principle:
Principle #2Taking out (Extraction)

2Adaptability or versatility

If audio feedback is emitted to enhance user experience and program immersion, then user engagement is improved, but interference with voice recognition algorithms increases

Engineering Contradiction:
Improveuser experienceVSAvoidcommand detection accuracy
Core Design Contradiction:
Adaptability or versatilityVSDifficulty of detecting and measuring

Solution Approach 1:

The system performs preliminary processing by pre-identifying and pre-removing program content audio from the environmental audio signal before the voice recognition algorithm processes the signal. This preliminary action prevents the recognition algorithm from encountering interfering audio patterns in the first place.

Inventive Principle:
Principle #10Preliminary action

3Productivity

If program audio content is played back with feedback loop, then audio system functionality is improved, but false command generation increases

Engineering Contradiction:
Improvesystem functionalityVSAvoidfalse commands
Core Design Contradiction:
ProductivityVSObject-generated harmful factors

Solution Approach 1:

The system converts the harmful feedback loop into a beneficial feature by using the deterministic delay characteristics to identify and remove program content audio. The feedback loop that previously caused false commands is now utilized to track and eliminate the interfering audio patterns, improving overall system reliability.

Inventive Principle:
Principle #22Blessing in disguise (Convert harm into benefit)

Data Source

PatentUS8738382B1Audio feedback time shift filter system and method
Publication Date: 2014.05.27 NVIDIA CORP
  • US8738382B1 patent drawing
  • US8738382B1 patent drawing
  • US8738382B1 patent drawing

AI summary

Audio feedback time shifted filtering systems and methods are presented. The systems and methods facilitate separation of program audio feedback from received environmental audio (e.g., audio sensed by a microphone.) The separation of the program audio feedback reduces interference from program content audio feedback on performance of voice recognition operations. In one embodiment of a personal video recorder audio filter method, environmental audio patterns are received, an audio feedback time shift filter process is executed for separating out program content from the environmental audio patterns, and voice recognition is performed on the filtered environment audio patterns (without interference from program audio content feedback). The time shift or deterministic delay provides a closer correlation between program audio content and program audio content feedback received at the microphone and permits input timing compensation to compensate for feedback loop delays.