Display Apparatus Multi-Language Voice Recognition Control

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing display apparatuses face limitations in voice recognition control when the system language differs from the language of the displayed content, preventing users from selecting hyperlinks or performing operations through voice commands.

Innovation Solution

A display apparatus and method that includes a processor to display text objects in a language different from the preset language, accompanied by a symbol, and performs operations based on voice recognition results, allowing users to select and interact with text objects across multiple languages by recognizing specific symbols or numbers.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If a voice recognition engine is determined in advance based on the system language, then the voice recognition system is simple to implement, but it cannot recognize voices in languages different from the system language

Engineering Contradiction:
Improvelanguage compatibilityVSAvoidvoice recognition system complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent introduces a language identification module as an intermediary between the voice recognition engine and the text output. This module detects the language of the displayed content and dynamically selects or configures the appropriate voice recognition engine, enabling multi-language support without requiring multiple pre-configured systems. The intermediary translates the user's voice in any language into the corresponding text object on the screen.

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The voice recognition system transitions from a static, pre-determined configuration to a dynamic system that adapts to the displayed content's language. The system automatically adjusts the voice recognition engine based on real-time detection of the content language, allowing flexible switching between different language modes without manual intervention or system reconfiguration.

Inventive Principle:
Principle #15Dynamics

2Ease of operation

If the system uses a fixed voice recognition engine for the system language, then the system is easy to control, but it fails to recognize hyperlink text in different languages through voice commands

Engineering Contradiction:
Improvevoice control capabilityVSAvoidvoice recognition accuracy
Core Design Contradiction:
Ease of operationVSReliability

Solution Approach 1:

The system implements a feedback mechanism where the language identification module continuously monitors the displayed content's language and provides feedback to the voice recognition module. This feedback loop ensures that the voice recognition engine is always configured to match the current content language, enabling accurate recognition of hyperlink text and other interactive elements regardless of the system's default language setting.

Inventive Principle:
Principle #23Feedback

Solution Approach 2:

The system performs preliminary language detection of the displayed content before processing voice commands. By identifying the language of the content in advance, the system pre-configures the appropriate voice recognition engine, ensuring that when the user speaks, the system is already prepared to accurately recognize and process the voice input in the correct language.

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS11726806B2Display apparatus and controlling method thereof
Publication Date: 2023.08.15 SAMSUNG ELECTRONICS CO LTD
  • US11726806B2 patent drawing
  • US11726806B2 patent drawing
  • US11726806B2 patent drawing

AI summary

A display apparatus is provided. The display apparatus according to an embodiment includes a display, and a processor configured to control the display to display a UI screen including a plurality of text objects, control the display to display a text object in a different language from a preset language among the plurality of text objects, along with a preset number, and in response to a recognition result of a voice uttered by a user including the displayed number, perform an operation relating to a text object corresponding to the displayed number.