Multilingual Speech Message Translation via Server Mediation
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Managers and employees are often isolated due to time constraints and language barriers, as they communicate across different devices and languages, necessitating improved multilingual asynchronous communication solutions.
Innovation Solution
A system and method for multilingual asynchronous communications that records speech messages, converts them to text, translates languages, and synthesizes speech in a target language, allowing for asynchronous communication without simultaneous availability and language compatibility.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If speech messages are transmitted across different languages and devices, then communication versatility is improved, but system complexity increases due to language translation and format conversion requirements
Solution Approach 1:
The patent introduces a message server as an intermediary component that coordinates communication between mobile devices, handles language translation requests, and manages format conversions. This mediator approach allows the system to achieve multilingual support without requiring direct integration between all device pairs, thereby reducing overall system complexity while maintaining high adaptability
Solution Approach 2:
The communications application is designed with multi-functional capabilities including speech recording, text conversion, language translation, and format adaptation within a single integrated system. This universal design allows the same application to handle multiple communication scenarios (different languages, devices, and formats) without requiring separate specialized systems for each case
2Ease of operation
If asynchronous communication is enabled without simultaneous availability requirements, then ease of operation is improved, but loss of time increases due to message conversion and translation processing delays
Solution Approach 1:
The system performs speech-to-text conversion and language translation in advance during message transmission and server processing, rather than delaying delivery until conversion is complete. This preliminary action allows the message to be prepared and staged for delivery, reducing the perceived wait time for the user while maintaining asynchronous communication benefits
3Adaptability or versatility
If speech messages are converted to text and translated across languages, then adaptability to different languages is improved, but device complexity increases due to multiple conversion processes
Solution Approach 1:
The message server acts as an intermediary that centralizes the speech-to-text and translation conversion processes, removing the burden of these complex operations from individual mobile devices. The device simply records speech and sends it to the server, which handles the conversion and returns the translated text, thereby achieving language adaptability without increasing device complexity
Solution Approach 2:
The patent extracts the complex speech-to-text and translation conversion functions from the mobile device and relocates them to the message server infrastructure. This extraction allows the device to remain simple while the server handles the computationally intensive conversion processes, achieving language adaptability without compromising device simplicity
Data Source
AI summary
Methods, systems, and computer program products are provided multilingual for asynchronous communications. Embodiments include recording a speech message in a digital media file; transmitting, from a sender multilingual communications application to a recipient multilingual communications application, the speech message in the digital media file; receiving, in the recipient multilingual communications application, the recorded speech message in the digital media file; converting, by the recipient multilingual communications application, the recorded speech message to text; identifying, by the recipient multilingual communications application, that the text of the recorded speech message is in a source language that is not a predetermined target language; translating, by the recipient multilingual communications application, the text in the source language to translated text in the target language; converting, by the recipient multilingual communications application, the translated text to synthesized speech in the target language; recording, by the recipient multilingual communications application, the synthesized speech in the target language in a digital media file; and playing the media file thereby rendering the synthesized speech.


