Adaptive STT and TTS Service Operation for Portable Terminals
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Portable terminals face challenges in adaptively operating Speech To Text (STT) and Text To Speech (TTS) services due to environmental and network limitations, leading to difficulties in maintaining effective communication services.
Innovation Solution
A system and method that allow for adaptive operation of STT and TTS services by converting speech signals to text and vice versa, using speech recognition devices and databases, and determining the appropriate communication format based on user preferences, terminal settings, and network conditions.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If the terminal continuously attempts call connection regardless of environment, then the user can maintain communication service availability, but the user experience deteriorates in unsuitable environments (conference rooms, bathrooms, libraries)
Solution Approach 1:
The system continuously monitors environmental context (location, noise level, network status) and uses this feedback to dynamically adjust communication service behavior. When the terminal detects it is in a conference room, bathroom, or library, it automatically switches to text-based communication or sends text messages, avoiding disruptive audio calls while maintaining service availability.
Solution Approach 2:
The communication service mode is made dynamic rather than static. The terminal automatically transitions between different communication modes (voice call, text call, text message) based on real-time environmental conditions, user context, and network status, allowing the service to adapt to changing situations without user intervention.
2Adaptability or versatility
If the terminal provides multiple communication service options (speech call, character call, image call), then the user has more communication flexibility, but the service complexity increases making it difficult for users to select and operate
Solution Approach 1:
The terminal automatically determines and selects the most appropriate communication service mode based on environmental context, network conditions, and user preferences, eliminating the need for users to manually evaluate and select from multiple options. The system serves itself by making intelligent decisions about which communication mode to use, reducing operational complexity while maintaining versatility.
Solution Approach 2:
The terminal integrates multiple communication functions (voice call, text call, image call, text messaging) into a single unified service framework. Rather than presenting separate services for users to choose from, the system provides a universal communication interface that automatically adapts its function based on conditions, combining multiple capabilities into one seamless service.
3Ease of operation
If the terminal automatically selects communication service based on environment, then the ease of operation improves, but the device complexity increases requiring context monitoring and adaptive control
Solution Approach 1:
The patent introduces a context management module that acts as an intermediary between the environment sensors and the communication service selection logic. This mediator collects and processes environmental context information (location, noise level, network status) and translates it into appropriate communication service decisions, isolating the complexity of adaptive control from the core communication functions and making the system more manageable.
Data Source
Figure 1
Figure 2
Figure 3
AI summary
An operation method capable of adaptively operating at least one of a Speech To Text (STT) service and a Text To Speech (TTS)service according to setting or user operation and a system thereof are provided. The method includes requesting a specific type of a communication service connection to a reception side terminal by a transmission side terminal, and performing an operation of at least one of a speech to text service providing speech recognition based text and a text to speech service converting the text into speech data between the reception side terminal and the transmission side terminal, and includes one of recognizing speech data provided from the transmission side terminal and converting the speech data into a text based on a first speech process supporting device connected to the transmission side terminal.