VR Speech Control System for Fatigue Reduction

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional virtual reality control methods require repetitive and inconvenient hardware operations, leading to user fatigue and a poor interactive experience.

Innovation Solution

A VR speech control method and apparatus that receives speech input, converts it to text, generates an intent object, and sends instructions to VR applications for execution, allowing users to control VR processes and scenarios through voice commands.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If conventional control methods (head rotation, gesture tracking) are used to control VR interactions, then the system can respond to user inputs, but the user experiences fatigue due to excessively repeated operations and poor interactive experience

Engineering Contradiction:
Improveease of operationVSAvoiddevice complexity
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The patent replaces mechanical control systems (head rotation, gesture tracking with wearable devices) with a voice-based control system. The speech recognition module converts user speech into control commands, eliminating the need for physical hardware interactions and reducing user fatigue while maintaining system responsiveness.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Solution Approach 2:

The patent introduces a speech recognition module as an intermediary between the user and the VR system. This mediator converts natural speech into structured control commands, providing a more natural and less fatiguing interaction method compared to direct mechanical control through head or hand movements.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Adaptability or versatility

If conventional control methods require repetitive hardware operations, then the system can achieve basic interaction functionality, but the interactive experience becomes inconvenient and causes user fatigue

Engineering Contradiction:
ImproveadaptabilityVSAvoidease of operation
Core Design Contradiction:
Adaptability or versatilityVSEase of operation

Solution Approach 1:

The patent implements a universal voice control interface that can handle multiple types of VR interactions through a single speech recognition system. The system adapts to different interaction scenarios (opening interfaces, controlling scenarios, operating objects) through natural language processing, providing versatile control without requiring different hardware for each function.

Inventive Principle:
Principle #6Universality (Multi-functionality)

3Ease of operation

If speech recognition is implemented in VR system, then user interaction convenience is improved, but system complexity increases due to additional processing modules

Engineering Contradiction:
Improveease of operationVSAvoiddevice complexity
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The speech recognition module serves as an intermediary layer that translates natural speech into structured control commands understood by the VR system. This mediator handles the complexity of speech processing internally while presenting a simple, natural interface to the user, effectively managing the trade-off between ease of operation and system complexity.

Inventive Principle:
Principle #24Intermediary (Mediator)

Data Source

PatentUS10714090B2Virtual reality speech control method and apparatus
Publication Date: 2020.07.14 BEIJING BAIDU NETCOM SCI & TECH CO LTD
  • US10714090B2 patent drawing
  • US10714090B2 patent drawing
  • US10714090B2 patent drawing

AI summary

A virtual reality speech control method and apparatus is provided in the present disclosure. The method includes: receiving a request for opening a speech interface from a VR application; opening a speech interface in response to the request and receiving input speech information through the speech interface; converting the speech information to text information, and normalizing the text information to generate an intent object in conformity with a preset specification; recognizing the intent object based on a preset specification set and acquiring an instruction corresponding to the intent object; sending the instruction to the VR application, such that the VR application executes the instruction and feeds back an execution result.