Speech Volume Feedback via Motion Object Display

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current speech recognition technologies do not effectively indicate whether a user's speech is being uttered at a volume sufficient for recognition, leading to potential misrecognition or noise interference.

Innovation Solution

An information processing device with a determination portion to assess user-uttered speech volume and a display controller that displays a moving object on a screen when the volume exceeds a recognizable threshold, allowing users to adjust their speech accordingly.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Loss of information

If speech recognition is performed without volume indication, then the system is simple, but the user cannot determine if speech is at recognizable volume

Engineering Contradiction:
Improvevolume informationVSAvoidsystem complexity
Core Design Contradiction:
Loss of informationVSDevice complexity

Solution Approach 1:

The patent implements a feedback mechanism by displaying a motion object that moves toward a display object when speech volume exceeds the recognizable threshold. This visual feedback provides users with real-time information about their speech volume, resolving the contradiction by adding necessary information without excessive complexity.

Inventive Principle:
Principle #23Feedback

Solution Approach 2:

The patent uses visual changes in the display object (motion objects moving toward the display object) to indicate speech volume status. This approach provides clear volume information through visual cues rather than complex auditory or text-based feedback systems.

Inventive Principle:
Principle #32Color changes

2Ease of operation

If no visual feedback is provided, then the display is simple, but users cannot adjust their speech volume appropriately

Engineering Contradiction:
Improvespeech adjustmentVSAvoiddisplay complexity
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The visual feedback system displays motion objects that move toward a display object based on speech volume, enabling users to intuitively understand and adjust their speech volume. This resolves the contradiction by providing essential operational guidance through a relatively simple visual mechanism.

Inventive Principle:
Principle #23Feedback

Solution Approach 2:

The display updates periodically or continuously based on speech input, providing dynamic visual feedback that guides users to adjust their speech volume. This periodic update mechanism delivers operational information without requiring complex continuous monitoring displays.

Inventive Principle:
Principle #19Periodic action

3Reliability

If speech volume threshold is not monitored, then processing is faster, but recognition accuracy decreases

Engineering Contradiction:
Improverecognition accuracyVSAvoidprocessing speed
Core Design Contradiction:
ReliabilityVSProductivity

Solution Approach 1:

The system performs preliminary volume assessment before full speech recognition processing by comparing speech volume against a predetermined threshold. This preliminary check ensures recognition accuracy by filtering out inaudible speech while maintaining processing efficiency through a simple threshold comparison rather than complex analysis.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent uses a fixed volume threshold parameter to determine whether speech is recognizable. This parameter-based approach maintains processing speed by using simple threshold comparison while ensuring reliability by establishing a clear criterion for acceptable speech volume.

Inventive Principle:
Principle #35Parameter changes

Data Source

PatentUS10642575B2Information processing device and method of information processing for notification of user speech received at speech recognizable volume levels
Publication Date: 2020.05.05 SONY GROUP CORP
  • US10642575B2 patent drawing
  • US10642575B2 patent drawing
  • US10642575B2 patent drawing

AI summary

To provide a technology capable of allowing a user to find whether speech is uttered with a volume at which speech recognition can be performed. Provided is an information processing device including: a determination portion configured to determine a user-uttered speech volume on the basis of input speech; and a display controller configured to control a display portion so that the display portion displays a display object. The display controller causes the display portion to display a first motion object moving toward the display object when the user-uttered speech volume exceeds a speech recognizable volume.