Artificial intelligence device for providing speech recognition function and method of operating artificial intelligence device

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing AI devices experience delays in responding to user commands due to the need to transmit all commands to a natural language processing server, especially when multiple wake-up words are used, leading to memory inefficiency and slowed speech recognition performance.

Innovation Solution

An AI device employing a basic wake-up word for activating speech recognition and an additional wake-up word for controlling operations, allowing the device to perform actions without server intervention in specific situations, thereby reducing the need for memory storage of all operation commands and enhancing speech recognition speed.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If all additional commands are transmitted to the NLP server for speech control, then comprehensive speech recognition is achieved, but delay occurs in data exchange process causing user inconvenience

Engineering Contradiction:
Improvespeech recognition accuracyVSAvoidresponse delay
Core Design Contradiction:
ReliabilityVSLoss of time

Solution Approach 1:

The patent segments wake-up words into two categories: basic wake-up words stored in cloud servers for comprehensive recognition, and additional wake-up words stored locally in the artificial intelligence device for immediate response. This segmentation allows the system to handle different types of commands through different paths, resolving the contradiction between comprehensive recognition and fast response.

Inventive Principle:
Principle #1Segmentation

2Adaptability or versatility

If multiple wake-up words are stored in the artificial intelligence device, then speech control versatility is improved, but memory capacity is limited or computation for recognizing wake-up word increases

Engineering Contradiction:
Improvewake-up word varietyVSAvoidmemory capacity requirement
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent divides wake-up word storage between cloud servers (basic wake-up words) and local device memory (additional wake-up words). This segmentation enables the device to support multiple wake-up words without overwhelming local memory capacity, while still providing comprehensive speech control capabilities.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent introduces cloud servers as an intermediary for storing and processing basic wake-up words, freeing up local device memory. The cloud server acts as a mediator that handles the storage burden, allowing the device to maintain versatility with limited local resources.

Inventive Principle:
Principle #24Intermediary (Mediator)

3Ease of operation

If basic wake-up word recognition activates speech recognition function, then speech control is enabled, but all commands require server transmission causing operational delays

Engineering Contradiction:
Improvespeech control activationVSAvoidcommand execution time
Core Design Contradiction:
Ease of operationVSLoss of time

Solution Approach 1:

The patent segments command processing into two paths: commands containing additional wake-up words are processed locally for immediate execution, while other commands are transmitted to the server. This segmentation enables fast local response for specific commands while maintaining comprehensive server processing for others.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent performs preliminary action by storing additional wake-up words locally in the device before they are needed. This pre-storation enables the device to immediately recognize and execute these wake-up words without waiting for server communication, reducing command execution time.

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS11200888B2Artificial intelligence device for providing speech recognition function and method of operating artificial intelligence device
Publication Date: 2021.12.14 LG ELECTRONICS INC
  • US11200888B2 patent drawing
  • US11200888B2 patent drawing
  • US11200888B2 patent drawing

AI summary

An artificial intelligence device for providing a speech recognition function includes a memory configured to store a basic wake-up word used to activate the speech recognition function of the artificial intelligence device and an additional wake-up word used to control operation of the artificial intelligence device, a microphone configured to receive a speech command, and a processor configured to determine whether a current situation is an additional wake-up word recognition situation when the basic wake-up word is recognized from the speech command and perform operation corresponding to the remaining command excluding the basic wake-up word from the speech command upon determining that the current situation is the additional wake-up word recognition situation.