Device-to-Device Communication Session Setup with Real-Time Transcription
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current communication systems lack efficient methods for users to connect with appropriate healthcare professionals, particularly in home settings, where users need to provide extensive information for each communication session, and there is a need for real-time transcription of audio for better understanding.
Innovation Solution
A device-to-device communication method that obtains user profile data to select suitable healthcare professionals, establishes communication sessions, and generates real-time transcripts of audio for users, reducing the need for additional input during session setup and enhancing understanding through simultaneous audio and text presentation.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If users manually provide information for each communication session in current systems, then communication sessions can be established, but the process requires extensive information input and is time-consuming
Solution Approach 1:
The system performs preliminary actions by obtaining and storing user profile data and selecting appropriate second devices in advance, before the actual communication session is requested. This allows the session setup to be expedited when the user initiates communication, as the matching and selection work has already been completed beforehand.
Solution Approach 2:
The system enables self-service by automatically selecting appropriate second devices based on stored profile data without requiring users to manually search or specify options. The system serves itself by maintaining profile information and performing automatic matching, reducing the burden on users during session setup.
2Ease of operation
If audio-only communication is used, then communication sessions can be established, but users may have difficulty understanding without visual support
Solution Approach 1:
The system merges audio communication with visual transcription by simultaneously delivering both the audio stream and its text transcript to the user's device. This combination allows users to benefit from both auditory and visual modalities, improving understanding while the system handles the complexity of coordinating both streams.
Solution Approach 2:
The system introduces text transcription as an intermediary element that bridges the gap between audio communication and user understanding. The transcription serves as a visual mediator that reinforces auditory information, making communication more accessible without requiring direct modification of the audio transmission itself.
3Extent of automation
If the system stores and processes user profile data, then appropriate second devices can be selected automatically, but the system complexity increases
Solution Approach 1:
The system implements a universal profile data structure that serves multiple functions: storing user preferences, identifying appropriate second devices, and enabling automatic matching. This multi-functional approach consolidates what could be separate complex systems into a single versatile framework, reducing overall system complexity while maintaining high automation.
Data Source
AI summary
A method of device to device communication is provided. The method may include obtaining a request for a communication session at a communication system from a first device. The method may further include obtaining profile data of a user associated with the first device and in response to obtaining the request, selecting a second device of multiple second devices to participate in the communication session using the profile data. In some embodiments, the method may also include establishing the communication session between the first device and the second device using the communication system and generating transcript data of the second device audio. The method may also include sending the transcript data and the second device audio to the first device for presentation of the transcription of the second device audio by the first device.


