Spatial Environmental Sound Text Display for Non-Obstructive Awareness
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing information presentation technologies do not effectively enhance usability by providing comprehensive and user-friendly presentation of surrounding environmental sounds and visual information to users, particularly for hearing-impaired individuals.
Innovation Solution
An information processing apparatus that estimates sound source positions and converts environmental sounds into text, calculates attention scores based on sound characteristics, recognizes objects from images, and generates presentation expression information to enhance usability by presenting relevant sound information without obstructing the user's view or behavior.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Loss of information
If environmental sound information is presented to the user, then the user's understanding of surrounding situation is improved, but the presentation may obstruct the user's view or behavior
Solution Approach 1:
The patent presents sound information in spatial dimensions by displaying text at positions corresponding to sound source locations in the visual field. This allows information to be added without blocking the user's forward view, as text is placed peripherally or to the sides rather than in the central visual path.
Solution Approach 2:
The patent applies different presentation strategies to different regions of the visual field. Text is displayed in peripheral areas where it provides information without obstructing central vision. The presentation position is dynamically adjusted based on the user's gaze direction and the location of important objects in the center of view.
2Loss of information
If comprehensive sound information is presented, then the user's situational awareness is improved, but the information overload may reduce usability
Solution Approach 1:
The patent extracts and prioritizes only the most relevant sound information for presentation. The calculation unit computes attention scores to identify which sounds warrant user attention, filtering out unnecessary information. This selective presentation of critical sound information maintains situational awareness without overwhelming the user.
Solution Approach 2:
The patent dynamically adjusts presentation parameters such as text size, color, and position based on the calculated attention scores. High-priority sounds receive more prominent visual treatment, while lower-priority sounds are minimized or omitted. This adaptive parameter adjustment optimizes information delivery based on current situational context.
Data Source
AI summary
An information processing apparatus according to an embodiment of the present technology includes an estimation unit, a first generation unit, a calculation unit, a recognition unit, and a second generation unit. The estimation unit estimates a position of a sound source on the basis of an environmental sound and a captured image around a user. The first generation unit generates text information obtained by converting the environmental sound into text. The calculation unit calculates an attention score regarding a risk level of the user on the basis of the text information. The recognition unit recognizes object information regarding an object on the basis of the captured image. The second generation unit generates presentation expression information to be presented to the user on the basis of the position of the sound source, the text information, the attention score, and the object information.


