Dynamic Visual Representation for Audio Streaming

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current social networking applications do not provide a seamless way for people to engage in meaningful live audio conversations beyond their immediate social network, limiting the expansion of communication and perspective sharing.

Innovation Solution

A method and system for initiating and streaming audio conversations using mobile applications, where users can select and connect with others, with visual representations of participants that change dynamically during speech, allowing for audio communication and profile management, enabling broader social interaction.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Loss of information

If video is used for audio conversations, then visual information is provided, but data usage and processing load increase

Engineering Contradiction:
Improvevisual informationVSAvoiddata usage
Core Design Contradiction:
Loss of informationVSUse of energy by moving object

Solution Approach 1:

The patent extracts only the essential visual elements (facial features) from complete video streams, transmitting merely the changed facial feature data rather than full video frames. This extraction approach provides necessary visual information while dramatically reducing data transmission requirements and processing energy consumption.

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The system applies different quality levels to different parts of visual information: high detail is provided only for facial features that change during speech, while the rest of the visual field uses lower resolution or static representations. This local quality differentiation optimizes the balance between visual information provision and data usage.

Inventive Principle:
Principle #3Local quality

2Ease of operation

If visual representations are generated in real-time, then user engagement is improved, but processing complexity increases

Engineering Contradiction:
Improveuser engagementVSAvoidprocessing complexity
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The system performs preliminary actions by pre-defining facial feature templates and change detection algorithms before real-time processing. Visual representations are generated by matching incoming video frames against pre-established templates, significantly reducing real-time processing complexity while maintaining engagement through dynamic visual feedback.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

Instead of generating entirely new visual representations in real-time, the system creates simplified copies or avatars that replicate essential facial feature movements. These copied representations provide sufficient visual engagement for user interaction while requiring far less computational complexity than full video processing.

Inventive Principle:
Principle #26Copying

3Loss of information

If dynamic facial features are tracked, then communication expressiveness is enhanced, but computational resources are consumed

Engineering Contradiction:
Improvefacial expression informationVSAvoidcomputational resources
Core Design Contradiction:
Loss of informationVSPower

Solution Approach 1:

The system extracts only the critical facial feature data necessary for expression communication, ignoring redundant visual information. By focusing solely on dynamic facial features that convey emotional and communicative content, the system enhances expressiveness while minimizing computational resource consumption.

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The patent monitors changes in facial feature parameters (position, shape, size) rather than processing entire facial images. This parameter-based approach tracks expressive movements efficiently, capturing essential communication information with significantly reduced computational requirements compared to full image analysis.

Inventive Principle:
Principle #35Parameter changes

Data Source

PatentUS11722328B2Complex computing network for improving streaming and establishment of communication among mobile computing devices based on generating visual representations for use in audio conversations
Publication Date: 2023.08.08 RIZZ IP LTD
  • US11722328B2 patent drawing
  • US11722328B2 patent drawing
  • US11722328B2 patent drawing

AI summary

Systems, methods, and computer program products are provided for improving establishment and broadcasting of communication among mobile computing devices based on generating visual representations for use in audio conversations.