Voice Recognition Apparatus Server Offload
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing voice recognition systems for home appliances face limitations in recognizing natural language voice commands due to resource constraints within individual apparatus, making it difficult to implement efficient and convenient control of home appliances across various languages.
Innovation Solution
A voice recognition apparatus and method that leverages a server system for natural-language voice recognition, utilizing a network infrastructure with technologies like Wi-Fi, Zigbee, and Bluetooth to communicate with a voice recognition server, allowing for the processing of voice commands without relying solely on local system resources.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Device complexity
If voice recognition is implemented using only local system resources in individual apparatus, then device complexity is reduced, but natural language recognition capability deteriorates due to computation requirements
Solution Approach 1:
The patent introduces a server as an intermediary component that handles complex natural language recognition tasks. The local apparatus (remote controller or home appliance) communicates with the server via network, delegating the computationally intensive voice recognition processing to the server while maintaining simple local hardware architecture.
2Ease of operation
If natural language voice recognition is implemented, then ease of operation is improved, but use of energy increases due to great amount of computation required
Solution Approach 1:
The server acts as an energy-efficient intermediary that performs computationally intensive natural language processing remotely. The local apparatus only needs to transmit voice data and receive commands, consuming minimal energy while still enabling sophisticated natural language recognition capabilities.
3Speed
If voice recognition processing is performed locally in each apparatus, then response speed is improved, but device complexity increases beyond embedded module capabilities
Solution Approach 1:
The system segments the voice recognition functionality into two parts: simple local components (microphone, transmitter, receiver) and complex processing (natural language recognition, command interpretation) that is performed remotely on the server. This segmentation allows fast local response for data transmission while offloading complex computation to the server.
Data Source
Figure 1~2
Figure 3
Figure 4(a)~4(b)
AI summary
Disclosed is a voice recognition apparatus including: an audio input unit configured to receive a voice; a communication module configured to transmit voice data received from the audio input unit to a server system, which performs voice recognition processing, and receive recognition result data on the voice data from the server system; and a controller configured to control the audio input unit and the communication module, wherein, when a voice command in the voice data corresponds to a pre-stored keyword command, the controller performs control to perform an operation corresponding to the keyword command, and wherein when the voice command in the voice data does not correspond to the pre-stored keyword command, the controller performs control to transmit the voice data including the voice command to the server system. Accordingly, voice recognition may be performed efficiently.