Multilingual Text Selection for Voice-Controlled Displays
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing display apparatuses face limitations in voice recognition control when the system language differs from the language used in hyperlink text, preventing accurate selection of hyperlinks or execution of commands.
Innovation Solution
A display apparatus and method that allows voice recognition across multiple languages by displaying text objects in a language different from the preset language with accompanying symbols or numbers, enabling operations based on voice recognition results, and utilizing a server for multi-language voice recognition.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If a voice recognition engine is determined in advance based on the system language, then the voice recognition system is simple to implement, but it cannot recognize voices in languages different from the system language
Solution Approach 1:
The voice recognition system dynamically selects the appropriate voice recognition engine based on the detected language of the hyperlink text, rather than using a fixed engine determined at system initialization. This allows the system to adapt to different languages while maintaining a modular architecture that manages complexity.
Solution Approach 2:
The system introduces an intermediary language detection mechanism that identifies the language of hyperlink text and selects the corresponding voice recognition engine. This intermediary layer enables multi-language support without requiring the system to maintain all possible voice recognition engines simultaneously active.
2Measurement precision
If the system uses a preset language for voice recognition, then the voice recognition accuracy is high for that language, but it fails to recognize hyperlink text in other languages
Solution Approach 1:
The system applies different voice recognition engines to different language contexts locally. Instead of using a single global voice recognition engine, it selects and applies the appropriate engine (e.g., Korean, English, Japanese) based on the specific language detected in the hyperlink text, ensuring high accuracy for each language while maintaining multi-language capability.
3Adaptability or versatility
If the display apparatus displays only text objects in the preset language, then the interface is simple to manage, but it cannot display text in other languages with proper voice control
Solution Approach 1:
The display interface dynamically adapts to show text objects in the language corresponding to the detected hyperlink text language, rather than being fixed to the system's preset language. This dynamic language adaptation enables multi-language display while the underlying modular architecture manages the complexity of handling multiple languages.
Data Source
AI summary
A display apparatus is provided. The display apparatus according to an embodiment includes a display, and a processor configured to control the display to display a UI screen including a plurality of text objects, control the display to display a text object in a different language from a preset language among the plurality of text objects, along with a preset number, and in response to a recognition result of a voice uttered by a user including the displayed number, perform an operation relating to a text object corresponding to the displayed number.


