Electronic Apparatus Speech Recognition Sensor Selection

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing electronic apparatuses face challenges in efficiently processing user speech across multiple sensor devices, leading to duplication of processing and resource wastage, particularly in distributed environments like living rooms and kitchens.

Innovation Solution

An electronic apparatus is designed to prioritize sensor devices by receiving audio signals from multiple devices, identifying the most effective sensor based on similarity and operation state, and performing speech recognition only with the identified device, thereby reducing redundant processing and resource utilization.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If multiple sensor devices receive and process user speech independently, then speech recognition coverage is improved, but processing duplication and resource wastage occur

Engineering Contradiction:
Improvespeech recognition coverageVSAvoidresource wastage
Core Design Contradiction:
Adaptability or versatilityVSLoss of energy

Solution Approach 1:

The system segments the speech processing task by designating one sensor device as the primary processor while others act as secondary devices that transmit audio data without performing full speech recognition, thereby dividing the computational workload and avoiding duplication

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The electronic apparatus acts as an intermediary that receives audio data from multiple sensor devices, determines which device should process the speech based on proximity and operational state, and routes the processing task accordingly to avoid redundant computation

Inventive Principle:
Principle #24Intermediary (Mediator)

2Measurement precision

If all sensor devices perform speech recognition, then recognition accuracy is improved, but network transmission and calculation resources are wasted

Engineering Contradiction:
Improverecognition accuracyVSAvoidnetwork and calculation resources
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The system extracts the speech recognition function from all sensor devices and concentrates it on the primary sensor device determined by proximity and operational state, while other devices retain only the audio data transmission function, thereby reducing overall system complexity

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The system performs preliminary determination of the primary sensor device based on proximity to the electronic apparatus and operational state before speech recognition occurs, pre-establishing which device will handle the computationally intensive recognition task

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS11417323B2Electronic apparatus and control method thereof
Publication Date: 2022.08.16 SAMSUNG ELECTRONICS CO LTD
  • US11417323B2 patent drawing
  • US11417323B2 patent drawing
  • US11417323B2 patent drawing

AI summary

An electronic device is provided. The electronic apparatus includes a communication interface, and at least one processor configured to receive a first audio signal and a second audio signal from a first sensor device, and a second sensor device located away from the first sensor device, respectively, through the communication interface, acquire similarity between the first audio signal and the second audio signal, acquire a first predicted audio component from the first audio signal based on an operation state of an electronic apparatus located adjacent to the first sensor device, and a second predicted audio component from the second audio signal based on an operation state of an electronic apparatus located adjacent to the second sensor device in the case where the similarity is equal to or higher than a threshold value, identify one of the first sensor device or the second sensor device as an effective sensor device based on the first predicted audio component and the second predicted audio component, and perform speech recognition with respect to an additional audio signal received from the effective sensor device.