Voice Recognition Search Space Segmentation for Accuracy

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

General voice recognition systems provide a universal single search space, making it difficult to offer customized services and resulting in lower recognition accuracy in specific situations.

Innovation Solution

Divide the search space into a general domain search space and a specific domain search space, allowing the voice recognition server to transition between them based on user-selected specific domains, creating and storing specific domain search spaces as needed for improved accuracy.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If a universal single search space is provided for all voice recognition service users, then the system structure is simple and easy to operate, but voice recognition accuracy is lowered under specific situations and customized services cannot be provided

Engineering Contradiction:
Improvesystem structure simplicityVSAvoidvoice recognition accuracy
Core Design Contradiction:
Ease of operationVSMeasurement precision

Solution Approach 1:

The search space is divided into multiple domains (e.g., general domain, specific domains like news, sports, entertainment) instead of using a single universal search space. Each domain has its own search space optimized for that particular type of content, allowing the system to select the appropriate domain-based search space for the specific voice recognition task to improve accuracy while maintaining operational simplicity through automated domain detection.

Inventive Principle:
Principle #1Segmentation

2Device complexity

If a universal single search space is provided for all voice recognition service users, then the device complexity is low, but the system cannot provide customized services suitable for different user situations

Engineering Contradiction:
Improvesearch space structureVSAvoidcustomized service capability
Core Design Contradiction:
Device complexityVSAdaptability or versatility

Solution Approach 1:

The system dynamically selects and switches between different domain-based search spaces based on the characteristics of the voice input and the recognition task. The search space structure is no longer static but adapts in real-time to the specific situation, allowing the system to provide customized services for different domains while keeping the overall device complexity manageable through modular architecture.

Inventive Principle:
Principle #15Dynamics

3Measurement precision

If multiple domain-specific search spaces are created and maintained, then voice recognition accuracy and customized service capability are improved, but the device complexity and storage requirements increase

Engineering Contradiction:
Improvevoice recognition accuracyVSAvoidsearch space management
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

Each domain-specific search space is optimized with local quality tailored to its specific domain characteristics. Instead of maintaining one generic search space with average performance across all domains, the system creates specialized search spaces with properties and parameters optimized for each domain (e.g., news domain search space has different characteristics than sports domain search space), thereby improving recognition accuracy for each specific domain while managing complexity through localized optimization.

Inventive Principle:
Principle #3Local quality

Data Source

PatentUS9520126B2Voice recognition system for replacing specific domain, mobile device and method thereof
Publication Date: 2016.12.13 ELECTRONICS & TELECOMM RES INST
  • US9520126B2 patent drawing
  • US9520126B2 patent drawing
  • US9520126B2 patent drawing

AI summary

A voice recognition system that divides a search space for voice recognition into a general domain search space and a specific domain search space. A mobile terminal receives a voice recognition target word from a user, and a voice recognition server divides a search space for voice recognition into a general domain search space and a specific domain search space and stores the spaces and performs voice recognition for the voice recognition target word through linkage of the general domain search space and the specific domain search space.