Voice Recognition Dictionary Creation for Multi-Language Support
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional voice recognition systems require specific acoustic models for each language and cannot automatically handle words in languages other than the recognized language, limiting their functionality to single-language support.
Innovation Solution
A recognition dictionary creation device that identifies the language of an input text, adds phonemes, and converts readings to match the phonemic system of the target language for voice recognition, enabling the creation of a recognition dictionary adaptable to multiple languages.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If a voice recognition system uses acoustic models for multiple languages, then it can support multiple languages, but it requires pre-specifying the language and adding readings manually, increasing device complexity
Solution Approach 1:
The system automatically identifies the language of input text and generates corresponding phoneme readings without requiring manual language specification or reading addition. The language identification unit detects the language, and the reading addition unit automatically creates phoneme sequences, enabling the system to serve itself rather than requiring user intervention for each new language input
Solution Approach 2:
The system pre-prepares phoneme readings by automatically converting text to phonemes before voice recognition occurs. The reading addition unit generates phoneme sequences in advance based on language identification, so that when voice recognition is performed, the phoneme data is already ready, eliminating the need for real-time language specification
2Ease of operation
If a voice recognition system automatically creates readings for target text, then it simplifies operation, but it can only handle the single language to be recognized
Solution Approach 1:
The reading addition unit is designed to handle multiple languages universally. It receives text in any language, identifies the language type, and automatically generates appropriate phoneme readings for that language. This makes the system capable of performing the same automatic reading creation function across multiple languages rather than being limited to a single language
3Adaptability or versatility
If the system converts readings from one phonemic system to another, then it enables multi-language recognition, but it increases processing time and complexity
Solution Approach 1:
The reading conversion unit performs phonemic system conversion in advance, before voice recognition occurs. By converting readings from the source language phonemic system to the target language phonemic system beforehand, the system prepares the data so that voice recognition can proceed directly without real-time conversion delays
Data Source
AI summary
A recognition dictionary creation device identifies the language of a reading of an inputted text which is a target to be registered and adds a reading with phonemes in the language identified thereby to the target text to be registered, and also converts the reading of the target text to be registered from the phonemes in the language identified thereby to phonemes in a language to be recognized which is handled in voice recognition to create a recognition dictionary in which the converted reading of the target text to be registered is registered.


