Voice Recognition Server Using Pre-stored Manuals for Display Control
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing voice recognition systems in display apparatuses struggle to accurately interpret complex spoken voices and provide appropriate responses, especially when user interactions require multiple steps or when hardware issues are detected, leading to incorrect operations or lack of response.
Innovation Solution
A voice recognition system that includes a server storing manuals for various display apparatuses, which recognizes spoken voices, generates response signals based on characteristic information, and processes operations accordingly, including guide messages or diagnosis results, and transmits these signals to the display apparatus for execution.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If the server analyzes the spoken voice signal to determine user intention and generates a response signal, then the voice recognition accuracy is improved, but the response time and system complexity increase
Solution Approach 1:
The system performs preliminary actions by storing multiple pre-analyzed response signals in the server before they are needed. When a user speaks a command, the server quickly matches the spoken voice against pre-prepared response signals rather than analyzing from scratch, significantly reducing response time while maintaining accuracy.
Solution Approach 2:
The voice recognition system is segmented into multiple components: voice input unit for capturing speech, server for storing and matching response signals, and processor for executing commands. This segmentation allows each component to specialize, improving overall recognition accuracy while distributing processing load to reduce response time.
2Ease of operation
If the system provides simple function execution for recognized voices, then the ease of operation is improved, but the adaptability to complex user interactions deteriorates
Solution Approach 1:
The server is designed with universal functionality to store and manage multiple types of response signals for various display apparatuses and functions. It can handle everything from simple channel changes to complex diagnostic procedures, making the system adaptable to diverse user needs while maintaining ease of voice-based operation.
Solution Approach 2:
The system dynamically adapts its response based on the complexity of the recognized voice command. For simple commands like 'channel-up', it executes directly. For complex queries like 'Tell me a recording method' or 'The screen is abnormal' It provides step-by-step guidance or initiates diagnostic routines, allowing the system to flexibly adjust its behavior to match user needs.
3Device complexity
If the server generates response signals without reflecting display apparatus characteristics, then the device complexity is reduced, but the reliability of the response deteriorates
Solution Approach 1:
The server stores response signals that are specifically tailored to different types of display apparatuses and their characteristics. Each response signal is customized for specific device types, ensuring that the guidance and commands are appropriate for the local characteristics of each apparatus, thereby improving reliability without requiring complex real-time analysis.
4Reliability
If hardware performance checking is added to respond to spoken voice, then the reliability is improved, but the device complexity increases
Solution Approach 1:
The system implements self-service diagnostics where the display apparatus automatically performs hardware performance checks when requested by voice command. The apparatus itself conducts the diagnosis and reports results to the server, eliminating the need for external diagnostic equipment and reducing overall system complexity while improving reliability.
Data Source
AI summary
A voice recognition system includes a server storing a plurality of manuals and a display apparatus transmitting, when a spoken voice of a user is recognized, characteristic information and a spoken voice signal corresponding to the spoken voice to the server, the characteristic information is characteristic information of the display apparatus, the server transmits a response signal to the spoken voice signal to the display apparatus based on a manual corresponding to the characteristic information among the plurality of manuals, and the display apparatus processes an operation corresponding to the received response signal; as a result, user convenience increases.


