360-Degree Video Speaker Identification via Text Labels
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional 360-degree video displays on electronic devices are limited by screen size, making it difficult for users to identify speakers located in different orientation regions, as the screen cannot display all orientation regions simultaneously, and users must manually navigate to find the speaker's screen when their voice is output.
Innovation Solution
An electronic device with a processor and memory that displays text corresponding to voices from different orientation regions, allowing users to select and view the screen of the speaker's orientation region, including features like speech bubbles and UI indicators to identify speakers.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If the display shows only one orientation region at a time, then the screen size limitation is respected, but the user cannot identify speakers from other orientation regions
Solution Approach 1:
The patent introduces text labels as an intermediary element that mediates between the visual display and the audio information. The text labels appear on the screen to indicate which speaker is currently speaking, allowing users to identify speakers from orientation regions not currently visible on the display without needing to manually navigate the 360-degree view.
2Loss of time
If the user manually navigates to find the speaker's screen, then the complete 360-degree view is accessible, but the time required to identify speakers increases
Solution Approach 1:
The patent applies preliminary action by displaying text labels that预先 (in advance) indicate which speaker is currently speaking before the user needs to locate them. This eliminates the need for users to manually scan through different orientation regions, as the text labels provide immediate information about the current speaker's location.
3Ease of operation
If the display shows all orientation regions simultaneously, then speaker identification is improved, but the screen size limitation cannot be overcome
Solution Approach 1:
The patent transitions from a two-dimensional display problem to a multi-dimensional solution by combining visual display with text labeling and audio information. Instead of trying to fit all orientation regions on a single 2D screen, the system uses text labels in the visual dimension plus spatial audio cues to provide comprehensive speaker identification across all orientation regions.
Data Source
AI summary
An electronic device is disclosed. In addition, various embodiments identified through the specification are possible. The electronic device includes a display, a processor, and a memory storing instructions that, when executed by the processor, cause the processor to display, when a video supporting a plurality of orientation regions is played, a screen of a first orientation region among the plurality of orientation regions and a first text corresponding to a voice of a first speaker in the screen, and display, in response to a user input of selecting a voice of a second speaker located in a second orientation region, a screen of the second orientation region.


