Smart Speaker Cognitive Sound Analysis and Response

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Smart speaker devices rely on fixed wake words or phrases for activation, limiting their ability to autonomously recognize different sounds and provide cognitive analysis of environmental events, failing to categorize sounds, identify patterns, and determine appropriate responses.

Innovation Solution

A smart speaker system capable of analyzing variable wake sounds, classifying sounds based on multiple characteristics, and determining responsive actions using a sound sample archive and cognitive analysis, allowing it to identify events and trigger appropriate responses without relying on fixed wake words.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If fixed wake words or phrases are used for smart speaker activation, then the device can reliably recognize when to activate, but it cannot autonomously recognize different sounds or provide cognitive analysis of environmental events

Engineering Contradiction:
Improvesound recognition capabilityVSAvoidautonomous sound classification
Core Design Contradiction:
Adaptability or versatilityVSExtent of automation

Solution Approach 1:

The system transitions from static wake word recognition to dynamic sound classification by continuously analyzing audio characteristics and adapting to different sound types in real-time, enabling the smart speaker to respond to various environmental events autonomously

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The smart speaker system performs self-service by autonomously classifying sounds, identifying events, and determining appropriate responses without requiring fixed wake words or human intervention, using its own audio capture and processing capabilities

Inventive Principle:
Principle #25Self-service

2Measurement precision

If the smart speaker system performs joint analysis of multiple sound characteristics, then sound classification accuracy is improved, but computational complexity increases

Engineering Contradiction:
Improvesound classification accuracyVSAvoidanalysis processing complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The sound analysis process is segmented into distinct stages: audio capture, characteristic extraction, joint analysis of multiple characteristics, and classification. This segmentation allows complex analysis to be broken down into manageable steps that can be processed efficiently

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

Sound models serve as intermediaries that bridge the gap between raw audio characteristics and classified sound types. These models pre-process and structure the analysis criteria, reducing the computational burden on the main processing system

Inventive Principle:
Principle #24Intermediary (Mediator)

3Adaptability or versatility

If the smart speaker system autonomously identifies events and triggers responsive actions, then functionality and environmental monitoring capability are enhanced, but system complexity increases

Engineering Contradiction:
Improveenvironmental monitoring capabilityVSAvoidsystem architecture complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The smart speaker system is designed with multi-functionality, combining audio capture, sound classification, event identification, and responsive action triggering within a single integrated system, eliminating the need for separate dedicated devices for each function

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

Sound models and event criteria are pre-configured and stored in the system, allowing for rapid classification and response without requiring complex real-time decision-making algorithms, thus reducing operational complexity

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS11631407B2Smart speaker system with cognitive sound analysis and response
Publication Date: 2023.04.18 INTERNATIONAL BUSINESS MACHINE CORPORATION
  • US11631407B2 patent drawing
  • US11631407B2 patent drawing
  • US11631407B2 patent drawing

AI summary

Smart speaker system mechanisms, associated with a smart speaker device comprising an audio capture device, are provided for processing audio sample data captured by the audio capture device. The mechanisms receive, from the audio capture device of the smart speaker device, an audio sample captured from a monitored environment. The mechanisms classify a sound in the audio sample data as a type of sound based on performing a joint analysis of a plurality of different characteristics of the sound and matching results of the joint analysis to criteria specified in a plurality of sound models. The mechanisms determine, based on the classification of the sound, whether a responsive action is to be performed based on the classification of the sound. In response to determining that a responsive action is to be performed, the mechanisms initiate performance of the responsive action by the smart speaker system.