Head-Mount Display Dynamic Caption Positioning for Sound Source Association
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing head-mount type display devices struggle to accurately associate sound sources with visual representations, as the fixed position of virtual images can hinder the visual field and fail to consider the relationship between sound sources and their corresponding visual cues, making it difficult for users to distinguish between multiple voices from different directions.
Innovation Solution
A transmissive display device with an image display section, sound acquisition section, conversion section, specific direction setting section, and display position setting section that allows users to set the position of character image light in their visual field based on the specific direction, ensuring the character image representing the sound is associated with the correct sound source and easily recognizable.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Loss of information
If the virtual image position is fixed in the visual field to allow the user to visually recognize the character image, then the user can see the character image representing the voice, but the character image may hinder the visual field of the user
Solution Approach 1:
The patent makes the character image display position dynamic by allowing the user to set a specific direction and automatically positioning the character image in that direction. This resolves the contradiction by making the display position adaptable rather than fixed, allowing the character image to be placed where it does not hinder the user's visual field while still being visible.
Solution Approach 2:
The patent introduces a directional dimension for character image placement by setting a specific direction in the visual field. Instead of using a fixed two-dimensional position, the system uses directional orientation to position character images, allowing them to be distributed across different directions and reducing visual field hindrance while maintaining visibility.
2Device complexity
If the character image is displayed at a fixed position, then the display structure is simple, but the relationship between the sound source and the character image is not considered
Solution Approach 1:
The patent implements feedback by using the sound acquisition section to detect the direction of the sound source and then automatically setting the specific direction for character image display based on that information. This creates a closed-loop system where the character image position is continuously adjusted to match the sound source direction, maintaining the association between sound and visual representation.
Solution Approach 2:
The patent introduces a specific direction setting section as an intermediary between the sound acquisition section and the character image display. This intermediary processes the sound source direction information and translates it into the appropriate display position, establishing the connection between sound source and character image while adding the necessary functional layer.
3Device complexity
If the character image is displayed without considering sound source direction, then the display process is simple, but it is difficult to distinguish multiple voices from different sound sources
Solution Approach 1:
The patent segments the visual field into different directional regions and assigns character images from different sound sources to different directions. By spatially separating character images based on their corresponding sound source directions, the system enables the user to distinguish between multiple voices without requiring complex processing of voice identities.
Solution Approach 2:
The patent uses visual differentiation in the form of directional positioning (analogous to color coding) to distinguish between multiple sound sources. Each sound source's character image is positioned in its specific direction, providing a visual cue that helps the user identify and distinguish between different voices in the environment.
Data Source
AI summary
A transmissive display device includes an image display section adapted to generate image light representing an image, allow a user to visually recognize the image light, and transmit an external sight, a sound acquisition section adapted to obtain a sound, a conversion section adapted to convert the sound into a character image expressing the sound as an image using characters, a specific direction setting section adapted to set a specific direction, and a display position setting section adapted to set an image display position, which is a position where character image light representing the character image is made to be visually recognized in a visual field of the user, based on the specific direction.


