Dual-Mode Voice Control System for Mobile Terminals

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing mobile terminals only support one voice input mode at a time, limiting user flexibility and convenience, as they do not allow selection or switching between 'operate-to-speak' and 'directly-speak' modes based on user behavior and microphone states.

Innovation Solution

A dual-mode voice control method and device that monitors user operations and microphone states to switch between 'operate-to-speak' and 'directly-speak' modes, allowing flexible selection and adaptation to different user habits by determining the start operation, monitoring time lengths, and switching voice modes based on preset thresholds.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If only one voice input mode is supported, then the system is simple and easy to implement, but user flexibility and convenience are limited

Engineering Contradiction:
Improvevoice mode flexibilityVSAvoidsystem complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The system dynamically switches between directly-speak mode and operate-to-speak mode based on real-time detection of user operations and microphone states. The voice input mode is not fixed but adapts automatically during operation, allowing the system to provide flexibility without requiring complex manual mode selection interfaces.

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The system automatically detects user intent through operation monitoring and microphone state detection, then autonomously selects the appropriate voice input mode. This self-service mechanism eliminates the need for explicit user mode selection, maintaining simplicity while achieving adaptability.

Inventive Principle:
Principle #25Self-service

2Reliability

If operate-to-speak mode is used when microphone is busy, then voice input can be captured, but mis-operations may occur due to insufficient user intent confirmation

Engineering Contradiction:
Improvevoice input accuracyVSAvoidoperation simplicity
Core Design Contradiction:
ReliabilityVSEase of operation

Solution Approach 1:

Before switching to operate-to-speak mode, the system performs preliminary detection of user operations (such as long-press duration, multiple taps) to confirm genuine user intent. This preliminary action prevents accidental mode switching and ensures that voice input is only activated when the user deliberately intends to speak.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The system provides feedback through visual indicators showing the current voice input mode and microphone state. This feedback mechanism helps users understand system status and prevents mis-operations by making the system state transparent and predictable.

Inventive Principle:
Principle #23Feedback

3Adaptability or versatility

If directly-speak mode is automatically selected, then the system responds quickly to user needs, but it cannot adapt to users who prefer operate-to-speak mode

Engineering Contradiction:
Improveuser preference adaptationVSAvoidmode switching time
Core Design Contradiction:
Adaptability or versatilityVSLoss of time

Solution Approach 1:

The system dynamically adapts to user preferences by monitoring operation patterns and automatically adjusting the default voice input mode. Over time, the system learns which mode each user prefers and switches accordingly, providing personalization without requiring explicit user configuration or mode selection time.

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The system automatically detects and adapts to user preferences through operation monitoring, eliminating the need for manual mode selection or configuration. This self-service adaptation occurs in real-time without interrupting user workflow or causing time loss.

Inventive Principle:
Principle #25Self-service

Data Source

PatentUS10373613B2Dual-mode voice control method, device, and user terminal
Publication Date: 2019.08.06 ALIBABA GROUP HOLDING LTD
  • US10373613B2 patent drawing
  • US10373613B2 patent drawing
  • US10373613B2 patent drawing

AI summary

A dual-mode voice control method is disclosed. The method may comprise determining whether a user has executed an operation of activating an operate-to-speak stop determination mode in a voice input interface. The method may further comprise, in response to determining that the user has executed the operation of activating the operate-to-speak stop determination mode, determining whether a microphone is in a busy state. The method may further comprise, in response to determining that the microphone is in the busy state, switching a voice mode from a directly-speak automatic stop determination mode to the operate-to-speak stop determination mode. Before the user executes the operation of activating the operate-to-speak stop determination mode, the voice mode is in the directly-speak automatic stop determination mode if the microphone is in the busy state.