Audio Message Pre-Conversion for Immediate Text Display

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current audio message processing methods require users to wait for a long time to receive converted text content, leading to increased anxiety and reduced communication efficiency, as they rely on real-time conversion of audio messages to text, which is inconvenient and inefficient, especially in situations where users cannot manually input data.

Innovation Solution

Implementing a method where a server proactively recognizes and converts audio messages to text in advance, allowing immediate display of text content on communication devices without the need for real-time conversion, even in situations without stable internet connectivity, by pre-fetching and pre-converting audio segments based on preset rules.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Loss of time

If real-time audio conversion is performed when user requests, then conversion accuracy is maintained, but user waiting time increases significantly

Engineering Contradiction:
Improveuser waiting timeVSAvoidcommunication efficiency
Core Design Contradiction:
Loss of timeVSProductivity

Solution Approach 1:

The system performs audio-to-text conversion in advance before the user actually needs to view or respond to the message. When an audio message is received, the device automatically converts it to text and stores both versions, so that when the user opens the message, the text is already available for immediate display, eliminating waiting time.

Inventive Principle:
Principle #10Preliminary action

2Speed

If audio messages are converted to text in advance, then response speed is improved, but device memory and processing resources are consumed

Engineering Contradiction:
Improveresponse speedVSAvoidmemory resource consumption
Core Design Contradiction:
SpeedVSQuantity of substance

Solution Approach 1:

The system selectively converts audio messages to text based on local conditions such as message importance, user preferences, and available resources. Not all audio messages are converted - only those that meet certain criteria are converted and stored, while others remain in original audio format, thus optimizing memory usage while maintaining fast response for critical messages.

Inventive Principle:
Principle #3Local quality

3Ease of operation

If manual text input is required, then communication precision is maintained, but user safety and convenience deteriorate in driving situations

Engineering Contradiction:
Improveuser convenienceVSAvoidcommunication safety
Core Design Contradiction:
Ease of operationVSReliability

Solution Approach 1:

The system replaces manual text input operations with automatic audio-to-text conversion. Users can send messages by speaking instead of typing, eliminating the need for manual keyboard interaction. This substitution maintains communication accuracy through voice recognition while dramatically improving safety for driving situations where manual input would be dangerous.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

4Ease of operation

If audio input functionality is added, then ease of operation is improved, but device complexity increases

Engineering Contradiction:
Improveinput convenienceVSAvoidsystem complexity
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The system integrates audio input functionality into the existing messaging framework, allowing the same communication application to handle both text and audio inputs through unified processing logic. The audio conversion feature shares resources with other messaging functions, such as using the same message queue, display interface, and notification system, thereby minimizing the increase in overall system complexity.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Data Source

PatentUS12046242B2Audio message processing method and apparatus
Publication Date: 2024.07.23 ALIBABA GROUP HOLDING LTD
  • US12046242B2 patent drawing
  • US12046242B2 patent drawing
  • US12046242B2 patent drawing

AI summary

Audio message processing methods and apparatuses are provided, where a method may include a server recognizing types of communication messages transmitted between communicating counterparties; when a type of any communication message is an audio type, the server acquiring the any communication message, and converting the any communication message to corresponding text content; and upon determining that any communicating party has a conversion need for the any communication message, the server sending the text content to the any communicating party. Through technical solutions of the present disclosure, text conversion may be performed upon audio messages in advance, thereby increasing response speed for audio conversion requests of users.