Voice Recognition Apparatus Server Offload

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing voice recognition systems for home appliances face limitations in recognizing natural language voice commands due to resource constraints within individual apparatus, making it difficult to implement efficient and convenient control of home appliances across various languages.

Innovation Solution

A voice recognition apparatus and method that leverages a server system for natural-language voice recognition, utilizing a network infrastructure with technologies like Wi-Fi, Zigbee, and Bluetooth to communicate with a voice recognition server, allowing for the processing of voice commands without relying solely on local system resources.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Device complexity

If voice recognition is implemented using only local system resources in individual apparatus, then device complexity is reduced, but natural language recognition capability deteriorates due to computation requirements

Engineering Contradiction:
Improvesystem resourcesVSAvoidnatural language recognition capability
Core Design Contradiction:
Device complexityVSAdaptability or versatility

Solution Approach 1:

The patent introduces a server as an intermediary component that handles complex natural language recognition tasks. The local apparatus (remote controller or home appliance) communicates with the server via network, delegating the computationally intensive voice recognition processing to the server while maintaining simple local hardware architecture.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Ease of operation

If natural language voice recognition is implemented, then ease of operation is improved, but use of energy increases due to great amount of computation required

Engineering Contradiction:
Improvevoice command convenienceVSAvoidcomputation energy
Core Design Contradiction:
Ease of operationVSUse of energy by moving object

Solution Approach 1:

The server acts as an energy-efficient intermediary that performs computationally intensive natural language processing remotely. The local apparatus only needs to transmit voice data and receive commands, consuming minimal energy while still enabling sophisticated natural language recognition capabilities.

Inventive Principle:
Principle #24Intermediary (Mediator)

3Speed

If voice recognition processing is performed locally in each apparatus, then response speed is improved, but device complexity increases beyond embedded module capabilities

Engineering Contradiction:
Improvevoice recognition responseVSAvoidcomputation resources
Core Design Contradiction:
SpeedVSDevice complexity

Solution Approach 1:

The system segments the voice recognition functionality into two parts: simple local components (microphone, transmitter, receiver) and complex processing (natural language recognition, command interpretation) that is performed remotely on the server. This segmentation allows fast local response for data transmission while offloading complex computation to the server.

Inventive Principle:
Principle #1Segmentation

Data Source

PatentEP3392878B1Voice recognition apparatus and voice recognition method
Publication Date: 2022.11.02 LG ELECTRONICS INC
  • EP3392878B1 patent drawingFigure 1~2
  • EP3392878B1 patent drawingFigure 3
  • EP3392878B1 patent drawingFigure 4(a)~4(b)

AI summary

Disclosed is a voice recognition apparatus including: an audio input unit configured to receive a voice; a communication module configured to transmit voice data received from the audio input unit to a server system, which performs voice recognition processing, and receive recognition result data on the voice data from the server system; and a controller configured to control the audio input unit and the communication module, wherein, when a voice command in the voice data corresponds to a pre-stored keyword command, the controller performs control to perform an operation corresponding to the keyword command, and wherein when the voice command in the voice data does not correspond to the pre-stored keyword command, the controller performs control to transmit the voice data including the voice command to the server system. Accordingly, voice recognition may be performed efficiently.