Multipoint Control Unit Multi-Language Audio Segmentation

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current video conference systems face challenges in managing multiple languages during conferences, leading to poor quality and inefficiency due to the need for multiple interpreters and the mixing of undesired languages, which disrupts the communication and pace of the conference.

Innovation Solution

A multipoint control unit employs multi-channel technology to process audio data from different language channels, allowing each conference site to receive and output speech in a selected language, using interpreting terminals to interpret and mix audio data independently for each language, thereby avoiding unnecessary language information and ensuring seamless communication.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If multiple interpreters are deployed at each conference site to handle multiple languages, then language communication quality is improved, but personnel costs and system complexity increase significantly

Engineering Contradiction:
Improvelanguage communication qualityVSAvoidsystem complexity
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

The patent segments the audio processing system into separate language channels. Each language has its own audio processing path, allowing independent handling of different languages without requiring physical separation of interpreters. The MCU divides audio data by language and processes each language stream independently through separate mixing channels.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent introduces an intermediary audio processing system between conference sites. Instead of requiring interpreters at each site, the MCU acts as an intermediary that receives audio from all sites, separates languages, processes them independently, and distributes appropriate language tracks to each conference site. This eliminates the need for human interpreters while maintaining multi-language communication quality.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Adaptability or versatility

If all conference sites speak in multiple languages simultaneously, then conference participation is improved, but audio mixing becomes chaotic and communication quality deteriorates

Engineering Contradiction:
Improveconference participationVSAvoidaudio mixing quality
Core Design Contradiction:
Adaptability or versatilityVSReliability

Solution Approach 1:

The patent segments the audio mixing process by language. The MCU creates separate mixing channels for each language, allowing multiple languages to be processed simultaneously without interference. Each conference site receives only the audio mix for its designated language, preventing the chaos of mixing multiple languages together while still allowing multi-language participation.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent applies local quality by providing language-specific audio processing at each conference site. Each site receives audio data tailored to its language requirements, and the audio characteristics are optimized for that specific language. This allows each participant to experience optimal audio quality for their language while others speak in different languages simultaneously.

Inventive Principle:
Principle #3Local quality

3Reliability

If interpreters are provided at each conference site, then language translation accuracy is improved, but conference efficiency and pace are reduced

Engineering Contradiction:
Improvelanguage translation accuracyVSAvoidconference efficiency
Core Design Contradiction:
ReliabilityVSProductivity

Solution Approach 1:

The patent uses the MCU as an intermediary that handles language translation and mixing centrally. Instead of having interpreters at each site who must individually translate and mix audio, the MCU processes all language translation and mixing operations automatically, maintaining accuracy while eliminating the time delays and coordination overhead of human interpreters.

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The patent enables continuous audio processing without interruption. The MCU continuously monitors audio from all conference sites, identifies language streams, processes translations, and mixes audio in real-time. This continuous automated processing eliminates the start-stop nature of human interpretation, maintaining conference pace and efficiency while ensuring accurate translation.

Inventive Principle:
Principle #20Continuity of useful action

4Adaptability or versatility

If audio data from multiple languages is mixed together, then comprehensive communication is achieved, but undesired language information disrupts participants

Engineering Contradiction:
Improvecommunication coverageVSAvoidlanguage interference
Core Design Contradiction:
Adaptability or versatilityVSObject-affected harmful factors

Solution Approach 1:

The patent segments audio data by language and transmits only the relevant language tracks to each conference site. The MCU separates mixed audio into distinct language channels and selectively distributes them, ensuring that participants receive only the language information appropriate for their site while blocking undesired language interference from other sites.

Inventive Principle:
Principle #1Segmentation

Data Source

PatentUS9031849B2System, method and multipoint control unit for providing multi-language conference
Publication Date: 2015.05.12 HUAWEI TECH CO LTD
  • US9031849B2 patent drawing
  • US9031849B2 patent drawing
  • US9031849B2 patent drawing

AI summary

A system for providing multi-language conference is provided. The system includes conference terminals and a multipoint control unit. The conference terminals are adapted to process a speech of a conference site, transmitting the processed speech to the multipoint control unit, process an audio data received from the multipoint control unit and output it. At least one of the conference terminals is an interpreting terminal adapted to interpret the speech of the conference according to the audio data transmitted from the multipoint control unit, process the interpreted audio data and output the processed audio data. The multipoint control unit is adapted to perform a sound mixing process of the audio data from the conference terminals in different sound channels according to language types, and then sends mixed audio data after the sound mixing process to the conference terminals.