Speech Codec Mode Grouping for Reduced Processing Complexity

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing speech encoding methods in cellular communication networks are complex and resource-intensive due to the need for accurate speech classification, which limits their usage and network capacity.

Innovation Solution

A method and apparatus for encoding frames using multiple codec modes with common parameter characteristics, where the codec mode selection is delayed to allow for more accurate parameter determination, reducing processing requirements and improving resource utilization.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If speech classification is performed using traditional variable rate encoding with SBRA algorithm, then speech quality is improved, but device complexity and processing resources increase significantly

Engineering Contradiction:
Improvespeech qualityVSAvoidprocessing complexity
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

The codec modes are divided into multiple groups, where each group contains codec modes with common parameter characteristics. This segmentation allows the encoder to select from pre-defined groups based on parameter matching rather than performing complex full-spectrum speech classification, thereby reducing processing complexity while maintaining speech quality.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The invention changes the approach from classifying speech signals to classifying codec modes based on their parameter characteristics. By grouping codec modes with common parameter characteristics and selecting groups based on parameter matching, the system reduces the computational burden of speech classification while preserving the ability to select appropriate encoding modes for different speech types.

Inventive Principle:
Principle #35Parameter changes

2Speed

If traditional speech encoding methods are used with early codec mode selection, then processing speed is improved, but measurement precision of speech parameters deteriorates

Engineering Contradiction:
Improveprocessing speedVSAvoidparameter determination accuracy
Core Design Contradiction:
SpeedVSMeasurement precision

Solution Approach 1:

Codec modes are pre-grouped according to their parameter characteristics before the encoding process. This preliminary organization allows for faster selection during encoding without requiring complex real-time analysis, as the encoder can directly match speech parameters to pre-defined groups rather than evaluating all possible codec modes.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The parameter groups serve as an intermediary layer between the speech signal and the codec modes. Instead of directly selecting codec modes from the full set based on speech analysis, the system uses parameter groups as a mediator that simplifies the selection process while maintaining accuracy in parameter determination.

Inventive Principle:
Principle #24Intermediary (Mediator)

3Adaptability or versatility

If multiple codec modes are available for selection, then adaptability is improved, but device complexity increases due to mode selection complexity

Engineering Contradiction:
Improvecodec mode adaptabilityVSAvoidmode selection complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The full set of codec modes is segmented into multiple groups, where each group contains codec modes with common parameter characteristics. This segmentation maintains the adaptability of having multiple codec modes available while reducing the complexity of selection by organizing modes into manageable groups that can be selected based on parameter matching rather than evaluating all modes individually.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

Each parameter group serves as a universal category that can accommodate multiple codec modes with similar characteristics. This multi-functionality allows the system to adapt to different speech types while using a unified grouping structure, reducing the need for separate selection logic for each codec mode.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Data Source

PatentUS7613606B2Speech codecs
Publication Date: 2009.11.03 NOKIA TECHNOLOGIES OY
  • US7613606B2 patent drawing
  • US7613606B2 patent drawing
  • US7613606B2 patent drawing

AI summary

A method of encoding a frame in a communication network using multiple codec modes, wherein the frame encoded by each codec mode is represented by multiple parameters. The method includes at least one stage, wherein the stage includes the steps of selecting one group from multiple groups of codec modes, wherein each group includes at least one codec mode and is arranged to have a common parameter characteristic. The method further includes encoding the frame with one of the codec modes from the selected group in dependence on the common parameter characteristic.