Occupant-Specific Audio Filtering in Vehicles
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Modern vehicles face challenges in distinguishing speech signals from multiple individuals speaking simultaneously or in rapid succession, which adversely affects speech recognition performance.
Innovation Solution
A method and system that utilize a position sensor to determine the positions and speech activity of occupants within a defined space, combined with microphones and processors to apply beamformers and time-frequency masks to separate audio signals and generate output signals corresponding to each occupant, effectively filtering sound and enhancing speech recognition.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If multiple microphones are used to receive sound from multiple occupants, then the ability to capture speech signals improves, but the difficulty of distinguishing and separating individual speech signals worsens
Solution Approach 1:
The patent segments the mixed audio signal into individual occupant speech signals by applying temporal-spatial filters to separate the speech of multiple occupants based on their spatial positions and temporal characteristics, resolving the difficulty of distinguishing individual speech signals
Solution Approach 2:
The patent introduces temporal-spatial filters as an intermediary processing layer between the microphones and speech recognition system, which mediates the separation of mixed speech signals by exploiting spatial and temporal differences in the audio signals from different occupants
2Adaptability or versatility
If speech recognition system processes mixed speech signals from multiple occupants, then the functionality for handling multiple speakers improves, but the recognition accuracy deteriorates
Solution Approach 1:
The patent segments the mixed speech signal into separate individual speech signals corresponding to each occupant before processing, allowing the speech recognition system to maintain high accuracy by processing separated signals rather than mixed signals
Solution Approach 2:
The patent performs preliminary separation of speech signals using temporal-spatial filtering before the speech recognition process, preparing distinct speech signals for each occupant in advance to ensure accurate recognition
Data Source
AI summary
Methods and systems are provided for filtering sound. A position sensor determines positions of a plurality of occupants in a defined space. Multiple microphones receive sound and generate corresponding audio signals. A processor in communication with the microphones and the position sensor receives the positions of the occupants and the audio signals. The processor determines which of the occupants are engaging in speech and applies a temporal-spatial filter to the audio signals to generate a plurality of output signals corresponding respectively to each occupant of the defined space.


