Speech Recognition Command Selection via Button Press Duration
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing speech recognition systems face challenges in efficiently executing multiple commands consecutively, as they often require cumbersome operations such as displaying lists and moving focus, making it difficult for users to select and confirm multiple recognition commands, especially when selecting media like videos or music.
Innovation Solution
An information processing apparatus that includes a first selection unit for speech recognition, a second selection unit for sequential command selection, a process determination unit to choose between these units based on button press duration, and an execution unit to execute the selected commands, allowing for single-operation sequential command selection and confirmation.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If speech recognition is used to select recognition commands, then the number of user input steps is reduced, but it becomes difficult to consecutively execute multiple commands when users need to select and confirm each content
Solution Approach 1:
The system dynamically switches between speech recognition mode and sequential selection mode based on the user's needs. The operation unit can detect different input patterns (speech vs. repeated button presses) and adaptively change the command selection method, allowing the system to be flexible in handling both single-command and multi-command execution scenarios
2Adaptability or versatility
If a list of recognition commands is displayed for selection, then users can select commands without speech recognition, but many operations are required including displaying the list, moving focus, and executing commands
Solution Approach 1:
The command selection process is segmented into two distinct methods: speech-based selection for single commands and sequential button-press-based selection for multiple commands. This segmentation allows each method to be optimized for its specific use case, reducing the operational burden for multi-command execution while maintaining flexibility
3Measurement precision
If users pronounce each content to select it sequentially, then accurate selection is achieved, but it becomes burdensome for users
Solution Approach 1:
The system provides a self-service mechanism where users can initiate sequential command execution through a simple button press pattern. Once triggered, the system automatically cycles through commands in a predetermined order, allowing users to simply observe and confirm the selection without needing to verbally pronounce each command, thereby reducing user burden while maintaining accuracy
Data Source
AI summary
An information processing apparatus performs a process in accordance with a command. The information processing apparatus includes a first selection unit configured to refer to a storage unit that stores a plurality of recognition commands for inputting the command by speech, recognize input speech and select a command based on the recognized input speech, and a second selection unit configured to sequentially select a plurality of commands that correspond to a plurality of recognition commands stored in the storage unit. The information processing apparatus further includes a process determination unit configured to select either the first selection unit or the second selection unit based on an operation performed on a predetermined operation unit, and an execution unit configured to execute a command which is selected by one of the selection units that is selected by the process determination unit.


