Voice Recognition Dictionary Creation for Multi-Language Support

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional voice recognition systems require specific acoustic models for each language and cannot automatically handle words in languages other than the recognized language, limiting their functionality to single-language support.

Innovation Solution

A recognition dictionary creation device that identifies the language of an input text, adds phonemes, and converts readings to match the phonemic system of the target language for voice recognition, enabling the creation of a recognition dictionary adaptable to multiple languages.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If a voice recognition system uses acoustic models for multiple languages, then it can support multiple languages, but it requires pre-specifying the language and adding readings manually, increasing device complexity

Engineering Contradiction:
Improvemulti-language supportVSAvoidlanguage specification and reading addition process
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The system automatically identifies the language of input text and generates corresponding phoneme readings without requiring manual language specification or reading addition. The language identification unit detects the language, and the reading addition unit automatically creates phoneme sequences, enabling the system to serve itself rather than requiring user intervention for each new language input

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The system pre-prepares phoneme readings by automatically converting text to phonemes before voice recognition occurs. The reading addition unit generates phoneme sequences in advance based on language identification, so that when voice recognition is performed, the phoneme data is already ready, eliminating the need for real-time language specification

Inventive Principle:
Principle #10Preliminary action

2Ease of operation

If a voice recognition system automatically creates readings for target text, then it simplifies operation, but it can only handle the single language to be recognized

Engineering Contradiction:
Improveautomatic reading creationVSAvoidlanguage support
Core Design Contradiction:
Ease of operationVSAdaptability or versatility

Solution Approach 1:

The reading addition unit is designed to handle multiple languages universally. It receives text in any language, identifies the language type, and automatically generates appropriate phoneme readings for that language. This makes the system capable of performing the same automatic reading creation function across multiple languages rather than being limited to a single language

Inventive Principle:
Principle #6Universality (Multi-functionality)

3Adaptability or versatility

If the system converts readings from one phonemic system to another, then it enables multi-language recognition, but it increases processing time and complexity

Engineering Contradiction:
Improvephonemic system conversion capabilityVSAvoidreading conversion process
Core Design Contradiction:
Adaptability or versatilityVSLoss of time

Solution Approach 1:

The reading conversion unit performs phonemic system conversion in advance, before voice recognition occurs. By converting readings from the source language phonemic system to the target language phonemic system beforehand, the system prepares the data so that voice recognition can proceed directly without real-time conversion delays

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS8868431B2Recognition dictionary creation device and voice recognition device
Publication Date: 2014.10.21 MITSUBISHI ELECTRIC MOBILITY CORP
  • US8868431B2 patent drawing
  • US8868431B2 patent drawing
  • US8868431B2 patent drawing

AI summary

A recognition dictionary creation device identifies the language of a reading of an inputted text which is a target to be registered and adds a reading with phonemes in the language identified thereby to the target text to be registered, and also converts the reading of the target text to be registered from the phonemes in the language identified thereby to phonemes in a language to be recognized which is handled in voice recognition to create a recognition dictionary in which the converted reading of the target text to be registered is registered.