Voice Personalization for Machine Reading via Sender Identification

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Machine reading technologies are impersonal and limited by a small set of available voices, lacking flexibility and intimacy in user interactions, particularly in situations where written communication is impractical.

Innovation Solution

Implementing a voice personalization system that allows users to select and authorize custom voice models for machine reading, enabling the use of personalized voices for various textual content, such as text messages, emails, and books, through a sender identification module, voice model module, and authorization module, which ensures secure and controlled usage.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If a generic voice model is used for machine reading, then the system is simple and easy to operate, but the user experience is impersonal and lacks flexibility

Engineering Contradiction:
Improvevoice personalizationVSAvoidsystem complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The system segments the voice model functionality into separate components: a generic voice model for basic operation and personalized voice models for enhanced user experience. Each voice model operates independently, allowing the system to switch between them based on user needs without increasing overall complexity

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

An intermediary module is introduced to manage voice model selection and authorization. This mediator handles the complexity of voice personalization by abstracting the underlying mechanisms, allowing users to benefit from personalized voices without directly interacting with the complex voice synthesis technology

Inventive Principle:
Principle #24Intermediary (Mediator)

2Adaptability or versatility

If custom voice models are allowed for each sender, then user intimacy and personalization are improved, but system complexity and authorization requirements increase

Engineering Contradiction:
Improvevoice selection flexibilityVSAvoidauthorization complexity
Core Design Contradiction:
Adaptability or versatilityVSEase of operation

Solution Approach 1:

The system implements self-service authorization where senders automatically grant permission for their voice to be used in text-to-speech conversion. The authorization is embedded in the message metadata, eliminating the need for manual user configuration and reducing operational complexity

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

Voice model authorization is performed in advance by the sender when creating their profile. This preliminary action stores the authorization status in the system, allowing rapid voice model selection without real-time authorization checks, thereby simplifying the user interaction

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS10176796B2Voice personalization for machine reading
Publication Date: 2019.01.08 INTEL CORP
  • US10176796B2 patent drawing
  • US10176796B2 patent drawing
  • US10176796B2 patent drawing

AI summary

Systems and techniques of voice personalization for machine reading are described herein. A message with textual content may be received. A sender of the message may be identified. A voice model that corresponds to the sender may be identified. An audio representation of the textual content may be rendered using the voice model.