Voice Operation Device Using Intermediary Recognition for Accuracy
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing voice operation systems face challenges in achieving high accuracy in voice recognition, leading to low effectiveness in operating devices as intended by users, due to limitations in voice data processing and lack of specialized terminology integration.
Innovation Solution
A voice operating system that includes a user terminal, a voice recognition device, and an operated device, where the user terminal receives voice input, performs voice recognition, and transmits the result to the operated device, utilizing a separate voice recognition device for accurate conversion and a natural language processor to execute commands, with the option for the user terminal to also serve as the voice recognition device, and downloading dictionary data for specialized terminology.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If voice data is transmitted directly to the operated device for recognition, then the system structure is simple, but the voice recognition accuracy is low
Solution Approach 1:
The patent introduces a voice recognition device as an intermediary component between the user terminal and the operated device. This separate voice recognition device processes the voice data and provides recognition results to the user terminal, which then transmits text data to the operated device. This intermediary structure resolves the contradiction by enabling high-accuracy voice recognition through specialized hardware while keeping the operated device's structure relatively simple.
2Reliability
If general voice recognition is used, then the system is easy to implement, but the effectiveness in operating devices as intended is low
Solution Approach 1:
The patent implements preliminary action by having the voice recognition device process voice data and generate text data before transmission to the operated device. The user terminal receives the text data, determines appropriate operation data based on this pre-processed information, and then transmits it to the operated device. This preliminary processing ensures that the operated device receives well-structured text data, improving operation effectiveness while maintaining reasonable implementation complexity.
3Object-affected harmful factors
If voice data is transmitted to the operated device, then all information is preserved, but user privacy is compromised
Solution Approach 1:
The patent extracts the voice recognition function from the operated device and places it in a separate voice recognition device. The user terminal receives voice input, the voice recognition device converts it to text data, and only this text data is transmitted to the operated device. This extraction principle protects user privacy by ensuring that voice data remains on the user terminal and is never transmitted to the operated device, while still providing the necessary information for device operation through the extracted text data.
Data Source
AI summary
A user terminal receives voice input from a user and transmits text data as a result of voice recognition performed by a voice recognition device on voice data indicating the received voice to an operated device, and the operated device operates in accordance with the text data received from the user terminal.


